Skip to content
Discussion options

You must be logged in to vote

The 429 is caused by the shape of the flow, not by OpenSearch. Agent 2 is effectively concatenating the repository into an LLM request, so the 1.35M requested TPM is expected. Repository ingestion should be a bounded ETL pipeline; the LLM should only see retrieved chunks at query time.

For OpenRAG 0.5.1, I would use this pipeline:

Git clone / GitHub Tree API
  -> deterministic file filtering
  -> one file per ingestion unit (bounded worker pool)
  -> Docling -> SplitText -> Embedding -> OpenSearch

Concretely:

  1. For a one-time import, clone/download the repository and use Knowledge -> Add Knowledge -> Folder. OpenRAG already processes folders in the background through the OpenSearch Ingestion

Replies: 1 comment

Comment options

You must be logged in to vote
0 replies
Answer selected by gabrielz06
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment
Category
Q&A
Labels
None yet
2 participants