We're hiring! Come build with us →
Zep

Context

The information environment around an agent: what it knows about users, the business and the world, and how that reaches the model.

63 posts

Governance and security

Defending Agent Memory Against Poisoning

One poisoned message, web page, or document can shape every later session that reads the same memory. Persistent memory makes prompt injection durable. We explain how memory poisoning works and which controls contain it, in your application and in Zep.

Engineering

Why we built a graph database service for agent memory

Konig, Zep's graph database service, is the data plane beneath our agent memory platform. This post covers why we built it, how it works, and what we learned along the way.

Integrations

Memory for every agent your team uses

A knowledge worker's agents don't share what they know: Claude and ChatGPT often can't reach company context, and the agents you build never see what Claude and ChatGPT learn. The Memory MCP Server gives each user one agent memory across all of them, governed by policy.

Integrations

Coding agents can design your Zep implementation

Zep now ships one plugin for Claude Code, Codex, and Cursor. It gives your coding agent the Zep documentation MCP server and the new building-with-zep skill, which encodes how to design and evaluate an agent memory implementation around the use case you need to deliver.

Agent memory

Evaluating Nemotron 3 Embed for agent memory

Zep's graph retrieval starts from entrypoints selected by semantic and BM25 search, so embedding recall bounds everything downstream. We benchmarked NVIDIA's Nemotron 3 Embed 1B, released today, on 5,954 production recall queries: first place against our production incumbent and two other models.

Governance and security

How Zep tracks provenance in agent memory

Agent memory is synthesized: an LLM derives facts from chat histories, documents, and business data. A derived fact matches no source word-for-word, so nothing ties it back to where it came from. Zep records that lineage; debugging, source-scoped retrieval, and compliance all depend on it.

Governance and security

Securing Agent Memory with ABAC

Attach policies to any Zep API key to control the exact endpoints it can call and the graph data it can read.

Integrations

Unified agent memory in any MCP client

Your team's agents are split across surfaces: the desktop assistant, the coding tool, the agents you build in-house. Each keeps its own memory or none at all. The Memory MCP Server puts them on one governed user graph, gated by your enterprise single sign-on.

Agent memory

Agent memory placement can cut your token bill up to 2x

Agent memory has to refresh every turn, but putting it in the system prompt breaks prompt caching and re-bills the whole conversation each turn. Here is the one-message fix, and the token savings it led to in an experiment.

Agent memory

Markdown is not agent memory

Some of the most capable agents in production keep their memory in plain markdown files, and for a single agent and a single user it is hard to beat. The pattern breaks in predictable places: at scale, as facts change and errors compound, and under concurrent agents.

Engineering

Building Agents in Go Without a Framework

A production agent is a long-running, concurrent, I/O-bound process that spends most of its time waiting on a model, a tool, or a human. That shape fits Go's runtime. This post explains why, surveys the Go framework options, and shows how to build an agent without one.

Benchmarks and evaluation

Sycophancy is a design choice

Writer's "Recalling Too Well" paper says memory systems amplify sycophancy. Its own data traces the amplification to two design decisions — one in Writer's experiment itself, one in a competitor's memory product.

Product updates

The Batch API: Load Large Datasets into Agent Memory

Zep's Batch API loads large datasets into agent memory faster, in batches up to 50,000 items, with a progress dashboard and no impact on real-time ingestion.

Retrieval

Smart Context Assembly: Fewer Tokens, Better Quality

Today we're announcing Smart Context Assembly, an upgrade to how Zep's default Context Block is built: higher accuracy from fewer tokens, with no code changes.

Context graphs

Observations: Patterns and Insights from the Context Graph

Observations are a new context type in Zep that capture patterns and insights across your Context Graphs, automatically discovered and surfaced to agents.

Context engineering

Zep's 5 Context Types: How to Use and Combine Each One

Zep produces five distinct types of context from a user's graph. Each captures something different. Here's when to reach for each, and how to combine them in one prompt.

Governance and security

Context You Can Trace, Filter, and Trust

Every fact in your agent's context graph came from somewhere. Zep's provenance architecture traces facts back to their source data — and lets you filter retrieval by origin.

Context engineering

Stop Letting Your Agent Decide What It Needs to Know

Your agent has the tools. It just doesn't call them — and smarter models won't fix that. Here's why the unknown unknowns problem is the hardest challenge in agent context, and what to do about it.

Context engineering

3 Decisions That Shape Every Agent's Context Architecture

Every agent context architecture comes down to three decisions: scope, data sources, and retrieval strategy. A framework for reasoning about persistent context for AI agents.

Context graphs

Build Better Context Graphs: Custom Instructions, Search Filters, and Webhooks

Custom extraction instructions, property-level search filters, exclusion filters, and webhooks — more control over how you build and query context graphs.

Context engineering

Context Templates: Context Engineering Made Simple

You're tuning an agent. Retrieve too little context and it hallucinates.

Benchmarks and evaluation

The Retrieval Tradeoff: What 50 Experiments Taught Us About Context Engineering

Zep builds temporal knowledge graphs from conversations and business data, then automatically retrieves relevant context when your agent needs it.

Context graphs

How Zep Works: A Visual Guide to Knowledge Graphs for AI Agents

Agents don't fail because the model is bad. They fail because they don't have the right context.

Integrations

Building Voice Agents with Memory: Zep x LiveKit

Create personalized voice agents with long-term memory with minimal added latency

Product updates

Agents That Always Remember What Matters

Zep now offers steerable user summaries. This post explores why these are useful and how best to implement them.

Integrations

Graphiti Hits 20K Stars! + MCP Server 1.0

Graphiti crossed 20,000 GitHub stars today! Thanks for building with us.

Engineering

How We Scaled Zep 30x in 2 Weeks (and Made It Faster)

Our infrastructure broke under 30x growth. Six weeks later, we made Zep faster than before: 10x better latency, 92% faster processing.

Context engineering

Zep v3: Context Engineering Takes Center Stage

Context engineering > prompt engineering: Zep v3 assembles memory & business data for agents that work.

Context engineering

What is Context Engineering, Anyway?

From Prompt Engineering to Context Engineering: The Why's and How.

Agent memory

The Private Agent Memory Fallacy

AI memory wallets sound appealing but face insurmountable economic, technical, and security challenges in practice.

Agent memory

Stop Using RAG for Agent Memory

Here's why you shouldn't be using RAG for agent memory—and what to do instead.

Product updates

Introducing Entity Types: Smarter, Structured Memory for Agents

Zep's new Entity Types let developers precisely structure and recall domain-specific information for more accurate, personalized agents.

Benchmarks and evaluation

Lies, Damn Lies, & Statistics: Is Mem0 Really SOTA in Agent Memory?

Mem0 claims State-of-the-Art in Agent Memory, but Zep outperforms it by 24%. We unpack why.

Benchmarks and evaluation

GPT-4.1 and o4-mini: Is OpenAI Overselling Long-Context?

We put OpenAI’s latest models through the LongMemEval benchmark—here’s why raw context size alone isn't enough.

Retrieval

The One-Token Trick

How single-token LLM requests can improve RAG search at minimal cost and latency.

Graphiti

Announcing a New Direction for Zep's Open Source Strategy

Ending support for Zep Community Edition to fully focus our open-source efforts on Graphiti

Integrations

Cursor IDE: Adding Memory With Graphiti MCP 🤖⚡️

Upgrade Cursor with persistent memory using Graphiti MCP. Now your favorite AI coding agent remembers your preferences, standards, and specs across sessions.

Product updates

Zep Q1 Product Round-up

Graph Explorer, an upgraded Playground, and a practical Cookbook—giving developers more control, better graph quality, and improved usability.

Integrations

Building a Memory Agent with the OpenAI Agents SDK and Zep

A video walkthrough demonstrating using Zep's agent memory and the new OpenAI Agents SDK to build an AI agent with long-term memory.

Graphiti

🎉 Big News! Graphiti is Launching on Product Hunt

Graphiti to launch on Product Hunt this Wednesday, February 19th!

Benchmarks and evaluation

Zep Is The New State of the Art In Agent Memory

Setting a new standard for agent memory with up to 100% accuracy gains and 90% lower latency.

Agent memory

Zep: A Temporal Knowledge Graph Architecture for Agent Memory

We introduce Zep, a novel memory layer service for AI agents that outperforms the current state-of-the-art systems.

Product updates

December 2024 Round-up

Explore User and Group Graphs, Fact Ratings, and SOC 2 Type 2 Certification!

Retrieval

How do you search a Knowledge Graph?

Can you build graph search that's both elegant & powerful? Here's how we did it in Graphiti, Zep's open source temporal Knowledge Graph library.

Context graphs

Building A Russian Election Interference Knowledge Graph

How we built a Knowledge Graph-based app for exploring Russian interference in the run-up to the 2024 US elections.

Context graphs

Exploring Russian Election Interference with a Knowledge Graph

We built a visual exploration tool and AI assistant for analyzing Russian interference in the run-up to the 2024 US elections.

Agent memory

Beyond Chat Memory: Making AI Interactions More Personal

Zep now connects user conversations and business data to help AI agents understand and serve users better.

Context graphs

Beyond Static Graphs: Engineering Evolving Relationships

Knowledge Graphs aren't adept at modeling changes in facts. This article explores the challenges we faced building time-aware Knowledge Graphs and our approaches to solving them.

Product updates

Announcing: Zep Community Edition

Zep Community Edition, open-sourced today, is the first Zep product powered by a Knowledge Graph, allowing you to build more accurate and personalized Agents.

Graphiti

Scaling LLM Data Extraction: Challenges, Design decisions, and Solutions

How we made building Knowledge Graphs faster and more dynamic

Graphiti

Graphiti: Temporal Knowledge Graphs for Agentic Apps

Graphiti builds dynamic, temporally aware knowledge graphs that represent complex, evolving relationships between entities over time.

Agent memory

Zep for Structured Outputs from Chat History (Video Walkthrough)

An end-to-end walkthrough demonstrating how to quickly and accurately extract data from chat histories stored in Zep.

Agent memory

Launching Structured Outputs from Chat History

Zep’s Structured Data Extraction is a high-accuracy tool for extracting data from chat histories. It's also 10x faster than gpt-4o.

Governance and security

Announcing Zep Archive: Records Retention and Right To Be Forgotten

Zep Simplifies LLM App Compliance with Data Privacy Regulations like GDPR and CCPA

Integrations

Foundations of LLM App Building in TypeScript

Learn how to build three foundational LLM apps using TypeScript, LangChain.js, and Zep.

Integrations

Zep ❤️ LlamaIndex: A Vector Store Walkthrough

LlamaIndex is a simple but powerful framework for building LLM apps. It's also an excellent tool for populating and searching Zep's Vector Store. This walkthrough demonstrates using LlamaIndex's new ZepVectorStore to do just that.

Product updates

Introducing the Zep Document Vector Store

With the addition of a Document Vector Store, Zep is now a single, batteries-included platform for grounding LLM apps with long-term memory.

Integrations

Diagnosing and Fixing Slow Chatbots with LangSmith and Zep

Poor chatbot response times can result in frustrated users and churn. Langchain’s new LangSmith service makes it easy to diagnose the cause of latency in an LLM app. In this article, we use LangSmith to analyze a very slow Langchain app and improve performance by an order of magnitude using Zep.

Governance and security

New Features: JWT Authentication, Azure OpenAI APIs, & Configurable Hard Deletion

Zep now supports JWT Authentication, Azure OpenAI APIs and OpenAI OrgIDs, and a configurable, periodic purge of soft-deleted data.

Agent memory

Personalizing LLM Interactions: Harnessing Generative Feedback Loops

LLM Applications can be personalized using Generative Feedback Loops through advanced memory & personalization.

Retrieval

Introducing Zep Hybrid Search and Custom Metadata

Zep now supports both vector search over message text and filtering on message metadata, including system metadata such as Named Entities and creation dates.

Integrations

LangchainJS Now Supports Zep!

LangchainJS now supports Zep Memory and Retrievers, allowing developers to take advantage of Zep's long-term memory, auto-summarization, vector search, and named entity extraction.

Agent memory

Introducing Zep: Long-term Memory Storage and Enrichment for AI Apps

Zep allows developers to focus on developing their AI apps, rather than building memory persistence, search, and enrichment infrastructure.

Sign up for Zep’s Newsletter

Writings on Zep, LLMs, and AI ecosystem tools.

No spam. Unsubscribe anytime.