Category
Topics
16 pages in this section.
Agent Evaluations
Agent Evaluations Overview Agent evaluations are the measurement layer for systems that plan, call tools, write code, retrieve context, or take actions over time. They combine offline tests, production traces, human revi
Agent Memory
Agent Memory Overview Agent memory is the set of mechanisms that lets an agent carry useful context across steps, sessions, users, repositories, documents, or decisions. It includes short term working context, long term
Agent Security
Agent Security Overview Agent security covers the controls that keep autonomous or semi autonomous AI systems within trusted boundaries. It includes authentication, authorization, tool permissions, sandboxing, prompt inj
Agent-Ready Accessibility
Agent Ready Accessibility Overview Agent ready accessibility is the overlap between accessible web engineering and agent operable web engineering. Both require reachable controls, explicit structure, understandable label
Agentic Search
Agentic Search Overview Agentic search is retrieval where an AI system actively plans, queries, follows leads, compares sources, and decides when it has enough evidence. It goes beyond one shot RAG by treating search as
Agentic Web
Agentic Web Agentic Web is the conference theme around making the public web usable by AI agents as an action surface, retrieval surface, interface layer, and data substrate. In this wiki it is a topic, not a standalone
AI Sandboxes
AI Sandboxes Overview AI sandboxes are controlled execution environments where agents can run code, browse, inspect files, call tools, or manipulate artifacts without putting the host system at unnecessary risk. A sandbo
Autoresearch
Autoresearch Overview AutoResearch is the use of agents to search, read, compare, synthesize, and sometimes design experiments over a body of evidence. The goal is not just summarization; it is repeatable research workfl
Coding Agents
Coding Agents Overview Coding agents are AI systems that can inspect repositories, reason about requirements, edit files, run commands, test changes, and sometimes open pull requests or operate development tools. They mo
Inference Engineering
Inference Engineering Overview Inference engineering is the practice of making AI model serving reliable, fast, cost aware, and fit for product constraints. It covers model selection, batching, caching, routing, quantiza
MCP Apps as Agentic App Runtime
MCP Apps as Agentic App Runtime Overview MCP Apps as an agentic app runtime is the idea that MCP servers can return interactive UI, not only structured text or JSON, so agents and humans can operate richer task surfaces
Model Context Protocol
Model Context Protocol Overview Model Context Protocol, or MCP, is a standard pattern for connecting AI applications to tools, data, and interactive capabilities through structured servers and clients. In this wiki it al
Nearly Headless Web
Nearly Headless Web Overview The nearly headless web is a product architecture where most discovery, understanding, and action can be handled by agents through machine operable surfaces, while selected moments still keep
Reachability Over Format
Reachability Over Format Overview Reachability over format is the idea that agent readiness depends less on inventing the perfect agent facing file format and more on making the right guidance reachable from the surfaces
Software Factories
Software Factories Overview Software factories are coordinated systems for turning ideas, issues, designs, tests, agents, and human review into shipped software. In an AI native version, multiple agents may handle planni
Voice Agents
Voice Agents Overview Voice agents are AI systems that understand, reason, and respond through speech, often in real time. They combine speech recognition, speaker diarization, language models, tool use, dialogue state,