-
-
Notifications
You must be signed in to change notification settings - Fork 8
Home
github-actions[bot] edited this page May 27, 2026
·
16 revisions
The Zero-Overhead, Agent-Ready AI Memory Backbone.
Welcome to the Spector Search documentation — your central hub for the high-performance, agent-native AI search engine. Whether you're connecting AI agents via MCP, building RAG pipelines, powering recommendation systems, or need sub-millisecond search with zero infrastructure, you're in the right place.
| Metric | Value |
|---|---|
| 🤖 MCP Tools | 6 agent-ready tools (semantic, hybrid, RAG, ingest, delete, status) |
| ⚡ Vector Search Latency | 0.05 ms avg @ 10K docs (128-dim) |
| 🔍 Keyword Search Latency | 0.98 ms avg @ 100K docs |
| 🧬 Hybrid Search Latency | 0.17 ms avg @ 10K docs |
| 🚀 Vector Throughput | 18,800 queries/sec @ 10K |
| 🧵 Concurrent Hybrid | 14,000+ ops/sec @ 16 threads (384-dim) |
| 🗜️ IVF-PQ + TurboQuant | 8–32× memory reduction |
| ✅ Test Suite | 331+ tests, all passing |
| 📦 Dependencies | Zero (JDK only) |
| Page | Description |
|---|---|
| Getting Started | Build, run, and search in 5 minutes |
| What is Spector Search | Product overview, use cases, and comparisons |
| JDK API Status | Vector API, Panama FFM, and preview feature compatibility |
| FAQ | Common questions answered |
| Page | Description |
|---|---|
| MCP Integration Architecture | How the MCP server works under the hood |
| MCP Server Guide | Setup for Claude Desktop, Cursor, and custom agents |
| Page | Description |
|---|---|
| Architecture Overview | Module diagram, data flow, threading model |
| Core Concepts | HNSW, IVF-PQ, BM25, RRF, SIMD deep-dives |
| Ingestion Pipeline | Document → chunk → embed → index pipeline |
| RAG Pipeline | End-to-end retrieval-augmented generation |
| Distributed Mode | Clustering, sharding, and replication |
| GPU Acceleration | CUDA setup and kernel details |
| Page | Description |
|---|---|
| REST API Reference | All endpoints with curl examples |
| Java SDK Guide | Programmatic usage (client + embedded) |
| Spring AI Integration | Spring AI VectorStore adapter |
| CLI Reference |
spectorctl commands |
| Configuration Guide | All parameters with tuning advice |
| Page | Description |
|---|---|
| Performance Tuning | Benchmarks and optimization strategies |
| Contributing | Development setup and PR process |
graph LR
A["🤖 AI Agent"] --> B["📡 MCP Server"]
B --> C["⚡ SpectorEngine"]
C --> D["🧠 Hybrid Search"]
D --> E["🎯 RRF Fusion"]
E --> F["🤖 LLM Re-ranking"]
F --> G["✨ Results"]
H["📄 Document"] --> I["🧩 Chunking"]
I --> J["🧬 Embedding"]
J --> C
Tip
New here? Start with Getting Started to build and run your first search in under 5 minutes. Want to connect an AI agent? See the MCP Server Guide.
| Language | Java 25 |
| License | Apache 2.0 |
| Modules | 18 Maven modules |
| Dependencies | Zero (JDK only) |
| SIMD | AVX2 / AVX-512 / NEON |
| GPU | CUDA via Panama FFM |
| MCP | Built-in, 6 agent-ready tools |
| Distributed | gRPC fan-out + consistent hashing |
Built with ⚡ by Spectrayan · GitHub · Apache 2.0 License
- Home
- Getting Started
-
Cognitive Memory
- Overview
- Getting Started
- Use Cases & Configuration
- API Reference
- Architecture
- The 6-Phase Scoring Pipeline
- Retrieval Stack
- Cognitive Profiles
- Salience & Importance
-
Biological Systems
- Overview
- Cortex — Tier Stores
- Hippocampus — Sleep Consolidation
- Synapse — Tags & Scoring
- Dopamine — Surprise Detection
- Amygdala — Emotional Valence
- 4-Layer Cognitive Graph
- Habituation — Anti-Filter Bubble
- Inhibition — Suppression
- Interference — Deduplication
- Prospective — Future Intents
- Metamemory — Self-Reflection
- Sync — Persistence & Replication
- Performance & Internals
- Cognitive Evaluation
- Synapse & Cortex
- Architecture
- Community