Daily micro briefs on AI agents and trust.
RAG-Safety-Bench is a benchmark for evaluating how retrieval-augmented generation systems handle safety risks when retrieving from document sources and responding to requests for harmful content.
The OpenAI Agents API is a product offering from OpenAI that enables developers to build and deploy agent systems.
LLM Visualizer is a tool or resource for understanding transformer architecture by building one from scratch.