FastRAGOSS
FeaturesHow it worksDevelopersDocsHosting
FeaturesHow it worksDevelopersDocsHosting
Overview

Getting started

  • Running locally
  • Deployment
  • Providers

Architecture

  • Architecture

Pipeline

  • Chunking
  • CRAG
  • Guardrails
  • Voice
  • Latency

Operations

  • Operations
  • Runbook
  • Benchmarking
  • LLM providers
documentation

Docs

Architecture, local setup, deployment, and pipeline deep-dives - rendered from the same markdown in the repo.

Getting started

Running locally

Compose stack, profiles, calibration, and first query.

Deployment

Vercel, Render, env vars, and cloud providers.

Providers

Local vs cloud adapter matrix.

Architecture

Architecture

Request path, components, and data flow.

Pipeline

Chunking

Strategies, metadata, and index shape.

CRAG

Corrective retrieval grading and rewrite.

Guardrails

Input gates before retrieval and generation.

Voice

STT, streaming voice queries, and audio handling.

Latency

Stage budgets and the sub-200ms retrieval target.

Operations

Operations

Monitoring, tracing, and day-two tasks.

Runbook

Incident checks and recovery steps.

Benchmarking

Latency harness and golden gates.

LLM providers

OpenAI-compatible generation backends.

FastRAGOSS

Voice-enabled, multilingual RAG. Answers only from retrieved sources - or it stays silent.

GitHubRohanVanshNitin

Product

  • Features
  • How it works
  • Hosting
  • Providers

Pipeline

  • Guardrails
  • Benchmarks
  • Latency
  • The team

Source

  • Repository
  • Docs
  • Running locally
  • Providers
  • Architecture

Team

  • Rohan Sen
  • Vansh Gularia
  • Nitin Pandey

2026 FastRAG. Open source.

github.com/rohansen856/FastRAG