@aiDotEngineer ยท English

๐Ÿค– AI Engineer

AI Engineer conference talks, workshops, and panels. Frontier labs, agents, RAG, evals, the AI engineering toolchain.

1107 videos ยท 20 topics ยท Browse by category โ†’ ยท YouTube โ†— ยท NotebookLM โ†—

Topics

๐Ÿค– Agents 175

LLM agents that plan, call tools, and act in loops. LangGraph, CrewAI, AutoGen, custom orchestration, multi-agent systems, agent reliability.

agentsmulti-agentmcpworkflows

๐Ÿ’ป Code Generation 130

AI coding tools and agents. Cursor, Devin, Copilot internals, SWE-bench, agentic refactoring, repo-scale understanding.

coding-agentscode-generationai-codingcopilot

๐Ÿ“ Evals 105

How to actually measure LLM and agent quality โ€” golden sets, LLM-as-judge, regression gates, production tracing, observability.

evalsobservabilitybenchmarksllm-as-judge

๐Ÿ’ผ AI Business 86

Going to market with AI. Pricing, GTM, build-vs-buy, moats, enterprise adoption, vertical agents, ROI stories.

ai-businessenterprise-aistartupsforward-deployed-engineering

๐Ÿ—๏ธ AI Infrastructure 83

GPU clusters, training stacks, autoscaling inference, data pipelines, feature stores, observability for AI workloads.

agentsinfrastructureobservabilitydurable-execution

๐Ÿ”Ž RAG 66

Retrieval-augmented generation โ€” chunking, embeddings, hybrid search, rerankers, citation, evaluation. The dominant pattern for grounding LLMs in private data.

ragneo4jknowledge-graphsgraphrag

๐Ÿ’ฌ LLM Apps 64

End-to-end LLM-powered applications. Prompt + context plumbing, structured outputs, retry & repair, user feedback loops.

agentsllm-appsevalsrag

โœจ Product & UX 55

Designing AI features users actually want. Latency, trust, streaming, citations, undo, the "AI moment" in a product.

productai-uxdesignux

๐Ÿ›ก๏ธ Safety & Alignment 45

Prompt injection defenses, jailbreak resistance, hallucination mitigation, PII handling, red-teaming, responsible scaling.

securityagent-securityguardrailssafety

๐Ÿง  Foundation Models 43

Frontier LLM training, architecture choices, scaling, post-training (SFT/RLHF/DPO), evaluation, releases from OpenAI, Anthropic, Google, Meta, Mistral, etc.

foundation-modelsgeminiopen-modelsopen-weights

โšก Inference & Serving 43

Throughput and latency engineering. Continuous batching, paged attention, quantization, speculative decoding, vLLM/TensorRT/SGLang.

inferencequantizationopen-modelson-device

๐Ÿ”Œ MCP 42

Model Context Protocol โ€” how clients (Claude, Cursor, IDEs) connect to servers that expose tools, resources, and prompts.

mcpagentsanthropicenterprise

๐ŸŽ™๏ธ Voice 35

Real-time voice AI. ASR (Whisper), TTS, turn detection, latency, voice agents for phones, support, accessibility.

voicepipecatvoice-agentslatency

๐Ÿ› ๏ธ Tools & Frameworks 31

The AI engineering toolchain โ€” LangChain, LlamaIndex, DSPy, LangGraph, LangSmith, Braintrust, Inspect, AGENTS.md.

agentstypescriptcontext-engineeringdspy

๐ŸŽจ Multimodal 31

Vision-language models, video understanding, image generation, multimodal agents. GPT-4V, Claude vision, Gemini, open-source VLMs.

multimodalgeminivideo-generationgenerative-media

๐ŸŽฏ Fine-Tuning 28

Adapting pre-trained models โ€” full SFT, LoRA/QLoRA, DPO, preference tuning. When fine-tuning beats prompting + RAG.

fine-tuningreinforcement-learningpost-trainingrl

โœ๏ธ Prompt Engineering 14

Prompting patterns โ€” few-shot, chain-of-thought, ReAct, structured output, prompt management at scale.

prompt-engineeringevalsprompt-optimizationcontext-engineering

๐Ÿ”ฌ Research 13

Frontier research talks โ€” new architectures, training techniques, theoretical insights, paper deep-dives.

continual-learningreinforcement-learningagiexpertise

๐Ÿ“ฆ Misc 8

Talks that span multiple themes, panels, opening keynotes, and general AI Engineer content.

communityai-engineeringconferencerobotics

๐Ÿงฎ Embeddings & Vector DBs 7

Embedding models, chunking, hybrid retrieval, vector store choice (Pinecone, Qdrant, Weaviate, pgvector), reranking.

embeddingsvector-searchrecsysmultimodal