Skip to main content
Watch: Claude Opus 5.5 Built an Entire 3D World

RAG

14 items

9 posts, 5 tools

Blog
EmbeddingGemma 2 Guide: Local Multimodal Embeddings for Code

EmbeddingGemma 2 is a 740M open model that embeds code, images, video and audio in one space. Specs, Matryoshka storage math, and where it breaks.

Blog
SearchOS Shows Deep Research Agents Need Shared State

SearchOS turns web research from a growing chat transcript into shared state: frontier tasks, evidence graphs, coverage maps, and failure memory. That is the pattern serious deep-research agents need.

Blog
Vector Database Comparison for RAG and AI Agents

pgvector, Pinecone, Qdrant, Weaviate, Chroma, Milvus, and Turbopuffer compared on hosting model, filtering, scale, and cost for RAG.

Blog
Agentic AI Reliability Is a Systems Problem

The Bayer and Thoughtworks PRINCE case study is a useful reminder that reliable agentic AI comes from context routing, traces, evals, monitoring, and human review, not from a better prompt alone.

Blog
Agentic Search Works Best When It Writes Queries, Not Answers

SNEWPAPERS is a useful Show HN signal: the strongest agentic search products do not replace search results with prose. They teach the agent to operate a real search system.

Blog
OpenAI Privacy Filter: Production PII Redaction Guide

OpenAI shipped an open-weight PII redactor. Here is how to wire it into a real ingestion pipeline locally, fast, with zero leaks, and how it benchmarks against Presidio and a regex baseline.

Blog
RAG with Claude: Add Context Without Retraining

A production-grade RAG pipeline with Claude. Chunking that survives real documents, retrieval tuning that actually moves the needle, citation tracking, and the prompt caching trick that makes RAG cheap enough to ship.

Tool
NotebookLM

Google's AI notebook that lets you ground a Gemini chat in your own uploaded sources. Generates summaries, mind maps, and podcast-style audio overviews.

Tool
Haystack

Open-source AI orchestration framework by deepset. Modular pipelines for RAG, agents, semantic search, and multimodal apps. Pipeline-as-graph architecture with explicit control.

Blog
AI Agent Memory Patterns

Agents forget everything between sessions. Here are the patterns that fix that: CLAUDE.md persistence, RAG retrieval, context compression, and conversation summarization.

Blog
What is RAG? Retrieval Augmented Generation Explained

How RAG works, why it matters, and how to implement it in TypeScript. The technique that lets AI models use your data without fine-tuning.

Tool
LangChain / LangGraph

Most popular LLM framework. 100K+ GitHub stars. Chains, RAG, vector stores, tool use. LangGraph adds stateful multi-agent workflows with cycles and persistence.

Tool
Mastra

TypeScript-first AI agent framework. Agents, tools, memory, workflows, RAG, evals, tracing, MCP, and production deployment for Node.js apps.

Tool
LlamaIndex

LLM data framework for connecting custom data sources to language models. Best-in-class RAG, data connectors, and query engines. Python and TypeScript.

AI Development Stack

Get Smarter About AI Dev

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.

One email per weekReal code, not theoryFree forever