Auteur • 198 livres
LlamaIndex Workflows : Data‑Connected Agents with Repeatable Retrieval Pipelines
Trex Team
Sandboxing AI Tools : Containers, MicroVMs, and Permission Scopes for Agents
Mastering the Model Context Protocol (MCP) : Standardizing Tool Access, Context, and Permissions for AI Agents
Vertex AI Agent Development Kit : From Prototype Agents to Governed Enterprise Deployments
OpenAI Agents SDK in Production : Architecting Reliable Tool‑Using Assistants
PydanticAI Cookbook : Typed Agents, Validated Outputs, and Schema‑First Reliability
AWS Bedrock AgentCore : Memory, Tools, and Event‑Driven Agents on AWS
BeeAI Framework : Building Agent Swarms with Open Protocols and Enterprise Controls
DSPy Prompt Programming : Data‑Driven Optimization for LLM Pipelines
Smolagents : Minimalist Agent Engineering with Tools, Code, and Guardrails
Haystack 2 Pipelines : Modular RAG, Agents, and Evaluation in Practice
Semantic Kernel Skills : Plugin‑First Copilots for Enterprise Apps
AutoGen Teams : Designing Multi‑Agent Collaboration Patterns That Don’t Collapse
CrewAI for Real Work : Role‑Based Agent Teams, Delegation, and Safe Autonomy
LangGraph : Graph‑Orchestrated Agents with Stateful, Testable Workflows
OpenHands : Building Task‑Executing Agents for Real Repos, Tickets, and Toolchains
OpenDevin Engineering : Self‑Hosted Software‑Dev Agents for Private Codebases
Continue.dev for Teams : Private Coding Assistants with Repo Context and Policy Controls
Aider in the Loop : Patch‑First AI Pair Programming with CI Safety Nets
Langfuse : Open-Source LLM Observability, Tracing, and Prompt Versioning
PromptLayer Ops : Managing Prompts, Experiments, and Releases Like Code
OpenLIT for GenAI : OpenTelemetry‑Style Observability for LLM Apps
Weave by W&B : End‑to‑End Experiment Tracking for LLM Products
Giskard for LLM QA : Detecting Harmful, Biased, and Broken Behaviors Before Launch
TruLens in Production : Feedback Functions, Scoring, and Continuous Evals
promptfoo in CI : Regression Testing Prompts, Tools, and RAG Pipelines
Phoenix for RAG Debugging : Traces, Retrieval Quality, and Hallucination Triage
Ragas for RAG : Measuring Retrieval, Faithfulness, and Answer Quality at Scale
DeepEval : Building an Automated LLM Evaluation Harness That Engineers Trust
Helicone Playbook : Monitoring, Caching, and Cost Controls for LLM APIs
LM Evaluation Harness : Measuring Model Quality with Reproducible Benchmarks
LiteLLM Proxy : Building a Multi‑Provider LLM Gateway with Routing and Budgets
Inspect AI : Writing Reproducible Evals and Safety Tests for LLM Systems
OpenAI Evals Cookbook : Designing Benchmarks for Product‑Grade LLM Features
OpenTelemetry for GenAI : Tracing Token Costs, Tool Calls, and RAG Latency
NeMo Guardrails : Policy‑Driven Safety for Tool‑Using Assistants
36 de 198 titres