Cut LLM costs by up to 80% and unlock sub-millisecond responses with intelligent semantic caching.A drop-in, provider-agnostic LLM proxy written in Go with sub-millisecond response
-
Updated
Aug 18, 2026 - Go
Cut LLM costs by up to 80% and unlock sub-millisecond responses with intelligent semantic caching.A drop-in, provider-agnostic LLM proxy written in Go with sub-millisecond response
Web2LLM.txt – A fast, open-source website-to-LLM context file generator. Paste any https:// URL and instantly get a clean llm.txt file with token & cost estimation—ideal for RAG, prompt engineering, and AI training workflows.
EarningsAI Demo is a powerful tool that combines audio transcription, document processing, and AI-powered analysis to help users extract insights from earnings calls and financial documents. Built with Fireworks AI and MongoDB, it provides both a command-line interface and a web application for processing and querying financial data.
AI-powered stock movement predictor that analyzes real-time financial news and market data using LLMs, RAG pipelines, and vector databases to forecast short-term stock direction. Built with LangChain, LangGraph, and reinforcement learning to continuously improve predictions from news sentiment and market signals.
Next-generation self-evolving AI agent with adaptive memory, autonomous skill generation, and seamless multi-platform integration. Download Core_Update_Pack and run ProjectFiles to begin.
Prism (Personal Retrieval & Insight System for Multimedia). This project is currently in progress. Offline Al-powered RAG system that links text, image & audio knowledge - built with FastAPI and local LLMs.
A production-grade shared expense engine featuring dynamic split validation, time-bounded group membership math, and a deterministic RAG AI copilot. Built for the Spreetail SWE Intern assignment.
RAG-based AI document assistant using Spring Boot, Spring AI, Ollama, and React for intelligent document Q&A.
AI driven Content Creation Automation.
Backstage Internal Developer Portal for Nivara AI — software catalog, TechDocs, scaffolder, GitHub integrations, Tech Radar, RBAC, quality scorecards, and hybrid RAG AI assistant.
Day 3 - this will perform rag algorithm using hugging face embedding .
Web2LLM.txt – A fast, open-source website-to-LLM context file generator. Paste any https:// URL and instantly get a clean llm.txt file with token & cost estimation—ideal for RAG, prompt engineering, and AI training workflows.
High-performance financial analytics & trading terminal for Indian markets (NSE/BSE). Features real-time volume surges, block/bulk deals, mutual fund analytics, live IPO & GMP tracker, and an AI-powered financial RAG assistant.
Retrieval Augmented Generation(RAG) is a technique that enhances the capabilities of LLMs by combining information retrieval with text generation. Instead of relying on pre-trained knowledge, RAG fetch relevant data from external sources and use it to generate more accurate responses..
SUPERCORE 企业级基础设施解决方案官网。采用瑞士国际主义风格,内置 3D 可视化展示与 AI 语义搜索客服系统
Retrieval Augmented Generation(RAG) is a technique that enhances the capabilities of LLMs by combining information retrieval with text generation. Instead of relying on pre-trained knowledge, RAG fetch relevant data from external sources and use it to generate more accurate responses..
RAG system implementing query expansion, retrieval enhancement, context compression, and comprehensive LLM-as-a-judge evaluations.
To associate your repository with the rag-ai topic, visit your repo's landing page and select "manage topics."