Rago

Rago is a lightweight framework for RAG.

Software License: BSD 3 Clause
Documentation: https://osl-incubator.github.io/rago

Features

Vector Database support
- FAISS
Retrieval features
- Support pdf extraction via langchain
Augmentation (Embedding + Vector Database Search)
- Support for Sentence Transformer (Hugging Face)
- Support for Open AI
- Support for SpaCy
Generation (LLM)
- Support for Hugging Face
- Support for llama (Huggin FAce)
- Support for OpenAI
- Support for Gemini

Roadmap

1. Add new Backends

As noted in several GitHub issues, our initial goal is to support as many backends as possible. This approach will provide valuable insights into user needs and inform the structure for the next phase.

2. Declarative API for Rago

Objective

To simplify and streamline the user experience in configuring RAG by introducing a declarative, composable API—similar to how Plotnine or Altair allows users to build visualizations.

Overview

The current procedural approach in Rago requires users to instantiate and connect individual components (retrieval, augmentation, generation, etc.) manually. This can become cumbersome as support for multiple backends grows. We propose a new declarative interface that lets users define their entire RAG steps in a single, fluent expression using operator overloading.

Proposed Syntax Example

from rago import Rago, Retrieval, Augmentation, Generation, DB, Cache

datasource = ...

rag = (
    Rago()
    + DB(backend="faiss", top_k=5)
    + Cache(backend="file")
    + Retrieval(backend="dummy")
    + Augmentation(backend="openai", model="text-embedding-3-small")
    + Generation(
        backend="openai",
        model="gpt-4o-mini",
        prompt_template="Question: {query}\nContext: {context}\nAnswer:"
    )
)

result = rag.run(query="What is the capital of France?", data=datasource)

Key Benefits

Intuitive Composition: Users can build complex pipelines by simply adding layers together.
Modularity: Each component is encapsulated, making it easy to swap or extend backends without altering the overall architecture.
Reduced Boilerplate: The declarative syntax minimizes the need for repetitive setup code, focusing on the "what" rather than the "how."
Enhanced Readability: The pipeline’s structure becomes immediately clear, promoting easier maintenance and collaboration.

Implementation Plan

Define Base Classes: Develop abstract base classes for each component (DB, Cache, Retrieval, Augmentation, Generation) to standardize interfaces and facilitate future extensions.
Operator Overloading: Implement the __add__ method in the main Rago class to allow chaining of components, effectively building the pipeline through a fluent interface.
Configuration and Defaults: Integrate sensible defaults and validation (using tools like Pydantic) so that users can override only when necessary.
Documentation and Examples: Provide comprehensive documentation and examples to illustrate the new declarative syntax and usage scenarios.

Installation

If you want to install it for cpu only, you can run:

$ pip install rago[cpu]

But, if you want to install it for gpu (cuda), you can run:

$ pip install rago[gpu]

Setup

Llama 3

In order to use a llama model, visit its page on huggingface and request your access in its form, for example: https://huggingface.co/meta-llama/Llama-3.2-1B.

After you are granted access to the desired model, you will be able to use it with Rago.

You will also need to provide a hugging face token in order to download the models locally, for example:

from rago import Rago
from rago.generation import LlamaGen
from rago.retrieval import StringRet
from rago.augmented import SentenceTransformerAug

# For Gated LLMs
HF_TOKEN = 'YOUR_HUGGING_FACE_TOKEN'

animals_data = [
    "The Blue Whale is the largest animal ever known to have existed, even "
    "bigger than the largest dinosaurs.",
    "The Peregrine Falcon is renowned as the fastest animal on the planet, "
    "capable of reaching speeds over 240 miles per hour.",
    "The Giant Panda is a bear species endemic to China, easily recognized by "
    "its distinctive black-and-white coat.",
    "The Cheetah is the world's fastest land animal, capable of sprinting at "
    "speeds up to 70 miles per hour in short bursts covering distances up to "
    "500 meters.",
    "The Komodo Dragon is the largest living species of lizard, found on "
    "several Indonesian islands, including its namesake, Komodo.",
]

rag = Rago(
    retrieval=StringRet(animals_data),
    augmented=SentenceTransformerAug(top_k=2),
    generation=LlamaGen(api_key=HF_TOKEN),
)
rag.prompt('What is the faster animal on Earth?')

Name	Name	Last commit message	Last commit date
Latest commit jgyasu feat: Add `microsoft/phi-2` for Text Generation (#91 ) Mar 26, 2025 9b2399f · Mar 26, 2025 History 74 Commits
.github	.github	chore(ci): Add ci workflow for closing stale prs (#84 )	Mar 18, 2025
artwork/logo	artwork/logo	fix: improve and fix documentation (#36 )	Feb 7, 2025
conda	conda	fix(pkg): Add support for MacOS (#75 )	Mar 13, 2025
docs	docs	docs(contributing): Fix minor syntax issue and clarify --no-verify us…	Mar 14, 2025
scripts	scripts	feat: Add initial structure for RAGO (#1 )	Oct 21, 2024
src/rago	src/rago	feat: Add `microsoft/phi-2` for Text Generation (#91 )	Mar 26, 2025
tests	tests	feat: Add `microsoft/phi-2` for Text Generation (#91 )	Mar 26, 2025
.editorconfig	.editorconfig	fix: improve and fix documentation (#36 )	Feb 7, 2025
.gitignore	.gitignore	fx: Fix undesired weekref logs attribute (#31 )	Jan 21, 2025
.makim.yaml	.makim.yaml	tests: Improve and re-organize tests and CI (#57 )	Mar 4, 2025
.pre-commit-config.yaml	.pre-commit-config.yaml	docs: Add information about roadmap (#81 )	Mar 17, 2025
.prettierignore	.prettierignore	docs: Add information about roadmap (#81 )	Mar 17, 2025
.prettierrc.yaml	.prettierrc.yaml	docs: Add information about roadmap (#81 )	Mar 17, 2025
.releaserc.json	.releaserc.json	fix: fix the semantic release workflow (#3 )	Oct 23, 2024
CODE_OF_CONDUCT.md	CODE_OF_CONDUCT.md	docs: Add information about roadmap (#81 )	Mar 17, 2025
LICENSE	LICENSE	feat: Add initial structure for RAGO (#1 )	Oct 21, 2024
README.md	README.md	docs: Add information about roadmap (#81 )	Mar 17, 2025
mkdocs.yaml	mkdocs.yaml	feat: Add initial structure for RAGO (#1 )	Oct 21, 2024
poetry.lock	poetry.lock	feat(augmented): add ChromaBD backend support (#54 )	Mar 25, 2025
pyproject.toml	pyproject.toml	feat(augmented): add ChromaBD backend support (#54 )	Mar 25, 2025

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

GitHub Sponsors

Repository files navigation

Rago

Features

Roadmap

1. Add new Backends

2. Declarative API for Rago

Objective

Overview

Proposed Syntax Example

Key Benefits

Implementation Plan

Installation

Setup

Llama 3

About

Releases 20

Sponsor this project

Contributors 6

Languages

License

osl-incubator/rago

Folders and files

Latest commit

History

Repository files navigation

Rago

Features

Roadmap

1. Add new Backends

2. Declarative API for Rago

Objective

Overview

Proposed Syntax Example

Key Benefits

Implementation Plan

Installation

Setup

Llama 3

About

Resources

License

Code of conduct

Stars

Watchers

Forks

Releases 20

Sponsor this project

Contributors 6

Languages