We removed 98.77% of an LLM’s input. Accuracy went up.
The obvious fix for an AI app missing information is to give it more context. But what happens when...
Tag archive
The obvious fix for an AI app missing information is to give it more context. But what happens when...
Archive777 isn't limited to conventional research questions. Some of the most interesting...
A six-week sweep of RAG/GraphRAG research: cheaper graph alternatives, a new GraphRAG index attack, and concrete vector-vs-graph guidance.

The previous article explained that RAG retrieves "semantically similar" chunks before generating an...
Large Language Models (LLMs) can answer many questions, but answering questions over a large and...
—————————————————————————————————————————— A retraction, cross-user RAG chains, and a...

Cohere, Perplexity and TopK all shipped embedding models in one week. Before you switch, build a tiny retrieval eval in TypeScript: answer recall, evidence recall, a ship gate, and a mixed-index check. No API key.

Everyone has been stuck with a support bot that keeps saying "Sorry, I didn't understand that" while...

You upload your company handbook, or a folder of PDFs, or two years of notes. You ask it something...

A legal team once sent my PDF-to-Markdown converter a 40-page services contract. Every desktop reader...

Ten vendor-neutral AI architectures as diagrams: RAG, agents with MCP, LLM gateway, guardrails, evals and more. Open them, zoom in, export them.
Calling an LLM API and wrapping a UI around it gives you a demo. Shipping it to real users takes a...