LlamaIndex
You will learn the data framework for RAG: ingestion and loaders, node parsing and chunking, indexes and vector stores, retrievers and query engines, agentic workflows, and evaluation. Interviewers ask because LlamaIndex is organised around the data pipeline, which forces you to hold opinions on chunking, index choice, and how retrieval quality gets measured.
part ofAI agent & RAG frameworksoverview, primer and where to startread it →on this pageshowhide
explore
- Data Ingestion & Loaders6 questions
- Node Parsing & Chunking6 questions
- Indexes & Vector Stores6 questions
- Retrievers & Query Engines6 questions
- Agents & Workflows6 questions
- Evaluation & Observability5 questions
questions
page 2 of 2When would you replace a LlamaIndex FunctionAgent with a hand-written Workflow?
basics
~20 sWhen the control flow is known in advance. A prebuilt agent pays an LLM turn to decide every step; a hand-written Workflow encodes the sequence in typed events and steps, making it cheaper, deterministic and testable — at the price of owning the loop yourself.
How would you prove a chunking change actually improved a LlamaIndex RAG pipeline?
basics
~20 sFreeze an eval set, then measure retrieval and generation separately: RetrieverEvaluator with hit_rate and MRR over labelled query/node pairs, and BatchEvalRunner with the faithfulness and correctness evaluators over answers. Hold the judge model fixed, and report cost and latency next to quality.
When should a LlamaIndex app move off SimpleVectorStore to Chroma, Pinecone or Weaviate?
basics
~20 sMove when the index stops fitting one process: the corpus exceeds memory, several replicas must share it, writes must be incremental rather than a whole-file rewrite, or durability and concurrent access become requirements. Until then SimpleVectorStore is faster to iterate on.
How do you decide between LlamaParse and LlamaIndex's built-in file readers for a PDF corpus?
basics
~20 sDecide by how much of the corpus is layout-dependent. Built-in readers extract a raw text stream and lose tables, column order and scanned pages; LlamaParse is a hosted parser returning structured markdown, at a per-page price, added latency, and data leaving your environment.
Your LlamaIndex query engine misses its p95 latency budget — which knobs do you trade?
basics
~20 sAttribute the time first: measure retriever.retrieve() alone against the full query() to see whether retrieval, postprocessing or synthesis dominates. Then trade deliberately — candidate pool size, reranker placement, response mode and streaming each buy latency back at a different quality cost.
showing 31–35 of 35