Autor*in • 180 Bücher
OpenAI Agents SDK in Production : Architecting Reliable Tool‑Using Assistants
Trex Team
Mastering the Model Context Protocol (MCP) : Standardizing Tool Access, Context, and Permissions for AI Agents
Vertex AI Agent Development Kit : From Prototype Agents to Governed Enterprise Deployments
PydanticAI Cookbook : Typed Agents, Validated Outputs, and Schema‑First Reliability
Smolagents : Minimalist Agent Engineering with Tools, Code, and Guardrails
Haystack 2 Pipelines : Modular RAG, Agents, and Evaluation in Practice
AWS Bedrock AgentCore : Memory, Tools, and Event‑Driven Agents on AWS
BeeAI Framework : Building Agent Swarms with Open Protocols and Enterprise Controls
DSPy Prompt Programming : Data‑Driven Optimization for LLM Pipelines
LlamaIndex Workflows : Data‑Connected Agents with Repeatable Retrieval Pipelines
Semantic Kernel Skills : Plugin‑First Copilots for Enterprise Apps
LangGraph : Graph‑Orchestrated Agents with Stateful, Testable Workflows
AutoGen Teams : Designing Multi‑Agent Collaboration Patterns That Don’t Collapse
CrewAI for Real Work : Role‑Based Agent Teams, Delegation, and Safe Autonomy
Aider in the Loop : Patch‑First AI Pair Programming with CI Safety Nets
Langfuse : Open-Source LLM Observability, Tracing, and Prompt Versioning
Continue.dev for Teams : Private Coding Assistants with Repo Context and Policy Controls
OpenHands : Building Task‑Executing Agents for Real Repos, Tickets, and Toolchains
OpenDevin Engineering : Self‑Hosted Software‑Dev Agents for Private Codebases
PromptLayer Ops : Managing Prompts, Experiments, and Releases Like Code
DeepEval : Building an Automated LLM Evaluation Harness That Engineers Trust
Weave by W&B : End‑to‑End Experiment Tracking for LLM Products
TruLens in Production : Feedback Functions, Scoring, and Continuous Evals
Giskard for LLM QA : Detecting Harmful, Biased, and Broken Behaviors Before Launch
promptfoo in CI : Regression Testing Prompts, Tools, and RAG Pipelines
Ragas for RAG : Measuring Retrieval, Faithfulness, and Answer Quality at Scale
Phoenix for RAG Debugging : Traces, Retrieval Quality, and Hallucination Triage
Helicone Playbook : Monitoring, Caching, and Cost Controls for LLM APIs
OpenLIT for GenAI : OpenTelemetry‑Style Observability for LLM Apps
Inspect AI : Writing Reproducible Evals and Safety Tests for LLM Systems
OpenTelemetry for GenAI : Tracing Token Costs, Tool Calls, and RAG Latency
OpenAI Evals Cookbook : Designing Benchmarks for Product‑Grade LLM Features
LM Evaluation Harness : Measuring Model Quality with Reproducible Benchmarks
LiteLLM Proxy : Building a Multi‑Provider LLM Gateway with Routing and Budgets
SPLADE Sparse Retrieval : Modern BM25‑Style Search for RAG Pipelines
ColBERT Reranking : High‑Recall Retrieval Without Blowing Your Latency Budget
36 von 180 Titel