rag-embedding-generation

Solid

Batch embedding generation with caching, rate limiting, and multiple provider support

AI & Automation 814 stars 53 forks Updated today MIT

Install

View on GitHub

Quality Score: 93/100

Stars 20%

Recency 20%

100

Frontmatter 20%

Documentation 15%

Issue Health 10%

License 10%

100

Description 5%

100

Skill Content

# RAG Embedding Generation Skill ## Capabilities - Generate embeddings with multiple providers - Implement batch processing for large datasets - Configure caching for embedding reuse - Handle rate limiting and retries - Support various embedding models - Implement embedding quality validation ## Target Processes - rag-pipeline-implementation - vector-database-setup ## Implementation Details ### Embedding Providers 1. **OpenAI Embeddings**: text-embedding-ada-002, text-embedding-3-* 2. **HuggingFace**: sentence-transformers models 3. **Cohere**: embed-v3 models 4. **Voyage AI**: voyage-2 models 5. **Local Models**: GGUF/ONNX embedding models ### Configuration Options - Model selection and parameters - Batch size optimization - Cache backend configuration - Rate limit settings - Retry policies - Dimensionality settings ### Best Practices - Use appropriate model for domain - Implement caching for cost reduction - Monitor embedding quality - Handle API errors gracefully ### Dependencies - langchain-openai / langchain-huggingface - numpy - Caching backend (Redis, SQLite)

Details

Author: a5c-ai
Repository: a5c-ai/babysitter
Created: 4 months ago
Last Updated: today
Language: JavaScript
License: MIT

Integrates with

OpenAI · AI Hugging Face · AI LangChain · AI Redis · Database SQLite · Database

Related Skills

AI & Automation Featured

videodb

See, Understand, Act on video and audio. See- ingest from local files, URLs, RTSP/live feeds, or live record desktop; return realtime context and playable stream links. Understand- extract frames, build visual/semantic/temporal indexes, and search moments with timestamps and auto-clips. Act- transcode and normalize (codec, fps, resolution, aspect ratio), perform timeline edits (subtitles, text/image overlays, branding, audio overlays, dubbing, translation), generate media assets (image, audio, video), and create real time alerts for events from live streams or desktop capture.

196,640 Updated 2 days ago

affaan-m

AI & Automation Featured

ck

Persistent per-project memory for Claude Code. Auto-loads project context on session start, tracks sessions with git activity, and writes to native memory. Commands run deterministic Node.js scripts — behavior is consistent across model versions.

196,640 Updated 2 days ago

affaan-m

AI & Automation Featured

browser

Web browser automation with AI-optimized snapshots for claude-flow agents

55,973 Updated today

ruvnet