brainy/README.md
David Snelling e311a149df docs: add deprecation warnings for addNoun and addVerb methods
- Add @deprecated JSDoc tags to TypeScript definitions
- Update all documentation examples to use modern add() and relate() API
- Preserve batch operations (addNouns, addVerbs) as they remain current
- Mark deprecated methods in both source and compiled definitions

Migration guide:
- addNoun(data, type, metadata) → add(data, { nounType: type, ...metadata })
- addVerb(source, target, type, metadata) → relate(source, target, type, metadata)
2025-09-17 10:48:41 -07:00

21 KiB
Raw Permalink Blame History

Brainy

Brainy Logo

npm version npm downloads MIT License TypeScript

🧠 Brainy 3.0 - Universal Knowledge Protocol™

World's first Triple Intelligence™ database unifying vector similarity, graph relationships, and document filtering in one magical API. Framework-friendly design works seamlessly with Next.js, React, Vue, Angular, and any modern JavaScript framework.

Why Brainy Leads: We're the first to solve the impossible—combining three different database paradigms (vector, graph, document) into one unified query interface. This breakthrough enables us to be the Universal Knowledge Protocol where all tools, augmentations, and AI models speak the same language.

Framework-first design. Built for modern web development with zero configuration and automatic framework compatibility. O(log n) performance, <10ms search latency, production-ready.

🎉 What's New in 3.0

🧠 Triple Intelligence™ Engine

  • Vector Search: HNSW-powered semantic similarity
  • Graph Relationships: Navigate connected knowledge
  • Document Filtering: MongoDB-style metadata queries
  • Unified API: All three in a single query interface

🎯 Clean API Design

  • Modern Syntax: brain.add(), brain.find(), brain.relate()
  • Type Safety: Full TypeScript integration
  • Zero Config: Works out of the box with memory storage
  • Consistent Parameters: Clean, predictable API surface

Performance & Reliability

  • <10ms Search: Fast semantic queries
  • 384D Vectors: Optimized embeddings (all-MiniLM-L6-v2)
  • Built-in Caching: Intelligent result caching
  • Production Ready: Thoroughly tested core functionality

Quick Start - Zero Configuration

npm install @soulcraft/brainy

🎯 True Zero Configuration

import {Brainy} from '@soulcraft/brainy'

// Just this - auto-detects everything!
const brain = new Brainy()
await brain.init()

// Add entities with automatic embedding
const jsId = await brain.add({
    data: "JavaScript is a programming language",
    type: "concept",
    metadata: {
        type: "language",
        year: 1995,
        paradigm: "multi-paradigm"
    }
})

const nodeId = await brain.add({
    data: "Node.js runtime environment",
    type: "concept",
    metadata: {
        type: "runtime",
        year: 2009,
        platform: "server-side"
    }
})

// Create relationships between entities
await brain.relate({
    from: nodeId,
    to: jsId,
    type: "executes",
    metadata: {
        since: 2009,
        performance: "high"
    }
})

// Natural language search with graph relationships
const results = await brain.find({query: "programming languages used by server runtimes"})

// Triple Intelligence: vector + metadata + relationships
const filtered = await brain.find({
    query: "JavaScript",                  // Vector similarity
    where: {type: "language"},           // Metadata filtering
    connected: {from: nodeId, depth: 1}  // Graph relationships
})

🌐 Framework Integration

Brainy 3.0 is framework-first! Works seamlessly with any modern JavaScript framework:

⚛️ React & Next.js

import { Brainy } from '@soulcraft/brainy'

function SearchComponent() {
  const [brain] = useState(() => new Brainy())

  useEffect(() => {
    brain.init()
  }, [])

  const handleSearch = async (query) => {
    const results = await brain.find(query)
    setResults(results)
  }
}

🟢 Vue.js & Nuxt.js

import { Brainy } from '@soulcraft/brainy'

export default {
  async mounted() {
    this.brain = new Brainy()
    await this.brain.init()
  },
  methods: {
    async search(query) {
      return await this.brain.find(query)
    }
  }
}

🅰️ Angular

import { Injectable } from '@angular/core'
import { Brainy } from '@soulcraft/brainy'

@Injectable({ providedIn: 'root' })
export class BrainyService {
  private brain = new Brainy()

  async init() {
    await this.brain.init()
  }

  async search(query: string) {
    return await this.brain.find(query)
  }
}

🔥 Other Frameworks

Brainy works with any framework that supports ES6 imports: Svelte, Solid.js, Qwik, Fresh, and more!

Framework Compatibility:

  • All modern bundlers (Webpack, Vite, Rollup, Parcel)
  • SSR/SSG (Next.js, Nuxt, SvelteKit, Astro)
  • Edge runtimes (Vercel Edge, Cloudflare Workers)
  • Browser and Node.js environments

📋 System Requirements

Node.js Version: 22 LTS or later (recommended)

  • Node.js 22 LTS - Fully supported and recommended for production
  • Node.js 20 LTS - Compatible (maintenance mode)
  • Node.js 24 - Not supported (known ONNX runtime compatibility issues)

Important: Brainy uses ONNX runtime for AI embeddings. Node.js 24 has known compatibility issues that cause crashes during inference operations. We recommend Node.js 22 LTS for maximum stability.

If using nvm: nvm use (we provide a .nvmrc file)

🚀 Key Features

World's First Triple Intelligence™ Engine

The breakthrough that enables the Universal Knowledge Protocol:

  • Vector Search: Semantic similarity with HNSW indexing
  • Graph Relationships: Navigate connected knowledge like Neo4j
  • Document Filtering: MongoDB-style queries with O(log n) performance
  • Unified in ONE API: No separate queries, no complex joins
  • First to solve this: Others do vector OR graph OR document—we do ALL

Universal Knowledge Protocol with Infinite Expressiveness

Enabled by Triple Intelligence, standardized for everyone:

  • 24 Noun Types × 40 Verb Types: 960 base combinations
  • ∞ Expressiveness: Unlimited metadata = model ANY data
  • One Language: All tools, augmentations, AI models speak the same types
  • Perfect Interoperability: Move data between any Brainy instance
  • No Schema Lock-in: Evolve without migrations

Natural Language Understanding

// Ask questions naturally
await brain.find("Show me recent React components with tests")
await brain.find("Popular JavaScript libraries similar to Vue")
await brain.find("Documentation about authentication from last month")

🎯 Zero Configuration Philosophy

Brainy 3.0 automatically configures everything:

import {Brainy} from '@soulcraft/brainy'

// 1. Pure zero-config - detects everything
const brain = new Brainy()

// 2. Custom configuration
const brain = new Brainy({
  storage: { type: 'memory' },
  embeddings: { model: 'all-MiniLM-L6-v2' },
  cache: { enabled: true, maxSize: 1000 }
})

// 3. Production configuration
const customBrain = new Brainy({
    mode: 'production',
    model: 'q8',           // Optimized model (99% accuracy, 75% smaller)
    storage: 'cloud',     // or 'memory', 'disk', 'auto'
    features: ['core', 'search', 'cache']
})

What's Auto-Detected:

  • Storage: S3/GCS/R2 → Filesystem → Memory (priority order)
  • Models: Always Q8 for optimal balance
  • Features: Minimal → Default → Full based on environment
  • Memory: Optimal cache sizes and batching
  • Performance: Threading, chunking, indexing strategies

Production Performance

  • 3ms average search - Lightning fast queries
  • 24MB memory footprint - Efficient resource usage
  • Worker-based embeddings - Non-blocking operations
  • Automatic caching - Intelligent result caching

🎛️ Advanced Configuration (When Needed)

Most users never need this - zero-config handles everything. For advanced use cases:

// Model is always Q8 for optimal performance
const brain = new Brainy()  // Uses Q8 automatically

// Storage control (auto-detected by default)
const memoryBrain = new Brainy({storage: 'memory'})     // RAM only
const diskBrain = new Brainy({storage: 'disk'})         // Local filesystem  
const cloudBrain = new Brainy({storage: 'cloud'})       // S3/GCS/R2

// Legacy full config (still supported)
const legacyBrain = new Brainy({
    storage: {forceMemoryStorage: true}
})

Model Details:

  • Q8: 33MB, 99% accuracy, 75% smaller than full precision
  • Fast loading and optimal memory usage
  • Perfect for all environments

Air-gap deployment:

npm run download-models        # Download Q8 model
npm run download-models:q8     # Download Q8 model

📚 Core API

search() - Vector Similarity

const results = await brain.search("machine learning", {
    limit: 10,                    // Number of results
    metadata: {type: "article"}, // Filter by metadata
    includeContent: true          // Include full content
})

find() - Natural Language Queries

// Simple natural language
const results = await brain.find("recent important documents")

// Structured query with Triple Intelligence
const results = await brain.find({
    like: "JavaScript",           // Vector similarity
    where: {                      // Metadata filters
        year: {greaterThan: 2020},
        important: true
    },
    related: {to: "React"}      // Graph relationships
})

CRUD Operations

// Create entities (nouns)
const id = await brain.add(data, { nounType: nounType, ...metadata })

// Create relationships (verbs)
const verbId = await brain.relate(sourceId, targetId, "relationType", {
    strength: 0.9,
    bidirectional: false
})

// Read
const item = await brain.getNoun(id)
const verb = await brain.getVerb(verbId)

// Update
await brain.updateNoun(id, newData, newMetadata)
await brain.updateVerb(verbId, newMetadata)

// Delete
await brain.deleteNoun(id)
await brain.deleteVerb(verbId)

// Bulk operations
await brain.import(arrayOfData)
const exported = await brain.export({format: 'json'})

🌐 Distributed System (NEW!)

Zero-Config Distributed Setup

// Single node (default)
const brain = new Brainy({
    storage: {type: 's3', options: {bucket: 'my-data'}}
})

// Distributed cluster - just add one flag!
const brain = new Brainy({
    storage: {type: 's3', options: {bucket: 'my-data'}},
    distributed: true  // That's it! Everything else is automatic
})

How It Works

  • Storage-Based Discovery: Nodes find each other via S3/GCS (no Consul/etcd!)
  • Automatic Sharding: Data distributed by content hash
  • Smart Query Planning: Queries routed to optimal shards
  • Live Rebalancing: Handles node joins/leaves automatically
  • Zero Downtime: Streaming shard migration

Real-World Example: Social Media Firehose

// Ingestion nodes (optimized for writes)
const ingestionNode = new Brainy({
    storage: {type: 's3', options: {bucket: 'social-data'}},
    distributed: true,
    writeOnly: true  // Optimized for high-throughput writes
})

// Process Bluesky firehose
blueskyStream.on('post', async (post) => {
    await ingestionNode.add(post, {
        nounType: 'social-post',
        platform: 'bluesky',
        author: post.author,
        timestamp: post.createdAt
    })
})

// Search nodes (optimized for queries)
const searchNode = new Brainy({
    storage: {type: 's3', options: {bucket: 'social-data'}},
    distributed: true,
    readOnly: true  // Optimized for fast queries
})

// Search across ALL data from ALL nodes
const trending = await searchNode.find('trending AI topics', {
    where: {timestamp: {greaterThan: Date.now() - 3600000}},
    limit: 100
})

Benefits Over Traditional Systems

Feature Traditional (Pinecone, Weaviate) Brainy Distributed
Setup Complex (k8s, operators) One flag: distributed: true
Coordination External (etcd, Consul) Built-in (via storage)
Minimum Nodes 3-5 for HA 1 (scale as needed)
Sharding Random Domain-aware
Query Planning Basic Triple Intelligence
Cost High (always-on clusters) Low (scale to zero)

🎯 Use Cases

Knowledge Management with Relationships

// Store documentation with rich relationships
const apiGuide = await brain.add("REST API Guide", {
    nounType: 'document',
    title: "API Guide",
    category: "documentation",
    version: "2.0"
})

const author = await brain.add("Jane Developer", {
    nounType: 'person',
    type: "person",
    role: "tech-lead"
})

const project = await brain.add("E-commerce Platform", {
    nounType: 'project',
    type: "project",
    status: "active"
})

// Create knowledge graph
await brain.relate(author, apiGuide, "authored", {
    date: "2024-03-15"
})
await brain.relate(apiGuide, project, "documents", {
    coverage: "complete"
})

// Query the knowledge graph naturally
const docs = await brain.find("documentation authored by tech leads for active projects")
// Find similar content
const similar = await brain.search(existingContent, {
    limit: 5,
    threshold: 0.8
})

AI Memory Layer with Context

// Store conversation with relationships
const userId = await brain.add("User 123", {
    nounType: 'user',
    type: "user",
    tier: "premium"
})

const messageId = await brain.add(userMessage, {
    nounType: 'message',
    type: "message",
    timestamp: Date.now(),
    session: "abc"
})

const topicId = await brain.add("Product Support", {
    nounType: 'topic',
    type: "topic",
    category: "support"
})

// Link conversation elements
await brain.relate(userId, messageId, "sent")
await brain.relate(messageId, topicId, "about")

// Retrieve context with relationships
const context = await brain.find({
    where: {type: "message"},
    connected: {from: userId, type: "sent"},
    like: "previous product issues"
})

💾 Storage Options

Brainy supports multiple storage backends:

// Memory (default for testing)
const brain = new Brainy({
    storage: {type: 'memory'}
})

// FileSystem (Node.js)
const brain = new Brainy({
    storage: {
        type: 'filesystem',
        path: './data'
    }
})

// Browser Storage (OPFS) - Works with frameworks
const brain = new Brainy({
    storage: {type: 'opfs'}  // Framework handles browser polyfills
})

// S3 Compatible (Production)
const brain = new Brainy({
    storage: {
        type: 's3',
        bucket: 'my-bucket',
        region: 'us-east-1'
    }
})

🛠️ CLI

Brainy includes a powerful CLI for testing and management:

# Install globally
npm install -g brainy

# Add data
brainy add "JavaScript is awesome" --metadata '{"type":"opinion"}'

# Search
brainy search "programming"

# Natural language find
brainy find "awesome programming languages"

# Interactive mode
brainy chat

# Export data
brainy export --format json > backup.json

🧠 Neural API - Advanced AI Features

Brainy includes a powerful Neural API for advanced semantic analysis:

Clustering & Analysis

// Access via brain.neural
const neural = brain.neural

// Automatic semantic clustering
const clusters = await neural.clusters()
// Returns groups of semantically similar items

// Cluster with options
const clusters = await neural.clusters({
    algorithm: 'kmeans',     // or 'hierarchical', 'sample'
    maxClusters: 5,          // Maximum number of clusters
    threshold: 0.8           // Similarity threshold
})

// Calculate similarity between any items
const similarity = await neural.similar('item1', 'item2')
// Returns 0-1 score

// Find nearest neighbors
const neighbors = await neural.neighbors('item-id', 10)

// Build semantic hierarchy
const hierarchy = await neural.hierarchy('item-id')

// Detect outliers
const outliers = await neural.outliers(0.3)

// Generate visualization data for D3/Cytoscape
const vizData = await neural.visualize({
    maxNodes: 100,
    dimensions: 3,
    algorithm: 'force'
})

Real-World Examples

// Group customer feedback into themes
const feedbackClusters = await neural.clusters()
for (const cluster of feedbackClusters) {
    console.log(`Theme: ${cluster.label}`)
    console.log(`Items: ${cluster.members.length}`)
}

// Find related documents
const docId = await brain.add("Machine learning guide", { nounType: 'document' })
const similar = await neural.neighbors(docId, 5)
// Returns 5 most similar documents

// Detect anomalies in data
const anomalies = await neural.outliers(0.2)
console.log(`Found ${anomalies.length} outliers`)

🔌 Augmentations

Extend Brainy with powerful augmentations:

# List available augmentations
brainy augment list

# Install an augmentation
brainy augment install explorer

# Connect to Brain Cloud
brainy cloud setup

🏢 Enterprise Features - Included for Everyone

Brainy includes enterprise-grade capabilities at no extra cost. No premium tiers, no paywalls.

  • Scales to 10M+ items with consistent 3ms search latency
  • Distributed architecture with sharding and replication
  • Read/write separation for horizontal scaling
  • Connection pooling and request deduplication
  • Built-in monitoring with metrics and health checks
  • Production ready with circuit breakers and backpressure

📖 Enterprise features coming in Brainy 3.1 - Stay tuned!

📊 Benchmarks

Operation Performance Memory
Initialize 450ms 24MB
Add Item 12ms +0.1MB
Vector Search (1k items) 3ms -
Metadata Filter (10k items) 0.8ms -
Natural Language Query 15ms -
Bulk Import (1000 items) 2.3s +8MB
Production Scale (10M items) 5.8ms 12GB

🔄 Migration from 2.x

Key changes for upgrading to 3.0:

  • Search methods consolidated into search() and find()
  • Result format now includes full objects with metadata
  • New natural language capabilities

🤝 Contributing

We welcome contributions! See CONTRIBUTING.md for guidelines.

🧠 The Universal Knowledge Protocol Explained

How We Achieved The Impossible

Triple Intelligence™ makes us the world's first to unify three database paradigms:

  1. Vector databases (Pinecone, Weaviate) - semantic similarity
  2. Graph databases (Neo4j, ArangoDB) - relationships
  3. Document databases (MongoDB, Elasticsearch) - metadata filtering

One API to rule them all. Others make you choose. We unified them.

The Math of Infinite Expressiveness

24 Nouns × 40 Verbs × ∞ Metadata × Triple Intelligence = Universal Protocol
  • 960 base combinations from standardized types
  • ∞ domain specificity via unlimited metadata
  • ∞ relationship depth via graph traversal
  • = Model ANYTHING: From quantum physics to social networks

Why This Changes Everything

Like HTTP for the web, Brainy for knowledge:

  • All augmentations compose perfectly - same noun-verb language
  • All AI models share knowledge - GPT, Claude, Llama all understand
  • All tools integrate seamlessly - no translation layers
  • All data flows freely - perfect portability

The Vision: One protocol. All knowledge. Every tool. Any AI.

Proven across industries: Healthcare, Finance, Manufacturing, Education, Legal, Retail, Government, and beyond.

→ See the Mathematical Proof & Full Taxonomy

📖 Documentation

Framework Integration

Core Documentation

🏢 Enterprise & Cloud

Brain Cloud - Managed Brainy with team sync, persistent memory, and enterprise connectors.

# Get started with free trial
brainy cloud setup

Visit soulcraft.com for more information.

📄 License

MIT © Brainy Contributors


Built with ❤️ by the Brainy community
Zero-Configuration AI Database with Triple Intelligence™