feat: implement clean embedding architecture with Q8/FP32 precision control
- Unified embedding system with single EmbeddingManager - Q8 model support with 75% smaller footprint (23MB vs 90MB) - Intelligent precision auto-selection based on environment - Clean cached embeddings with TTL and memory management - Zero-config setup with smart defaults - Complete storage structure documentation - Removed legacy worker and hybrid managers - Streamlined model configuration and precision management
This commit is contained in:
parent
3227ad907c
commit
184d5dcf34
23 changed files with 1575 additions and 1369 deletions
|
|
@ -13,6 +13,15 @@ export {
|
|||
logModelConfig
|
||||
} from './modelAutoConfig.js'
|
||||
|
||||
// Model precision manager
|
||||
export {
|
||||
ModelPrecisionManager,
|
||||
getModelPrecision,
|
||||
setModelPrecision,
|
||||
lockModelPrecision,
|
||||
validateModelPrecision
|
||||
} from './modelPrecisionManager.js'
|
||||
|
||||
// Storage configuration
|
||||
export {
|
||||
autoDetectStorage,
|
||||
|
|
|
|||
Loading…
Add table
Add a link
Reference in a new issue