AI Engineer World's Fair 2024
116 talks
Personality Driven Development: Exploring the Frontier of Agents with Attitude
Customized, production-ready inference with open source models
Optimizing LLMs in Insurance with DSPy
Claude Plays Minecraft
The Adversarial Path to the Personal Assistant
RAG at scale: production-ready GenAI apps with Azure AI Search
Training Albatross, An Expert Finance LLM
Accelerating Mixture of Experts Training With Rail-Optimized InfiniBand Networking in Crusoe Cloud
AI Music Generation, From Prompt to Production
System Design for Next-Gen Frontier Models
Giving a Voice to AI Agents
How to build the world's fastest voice bot
Building an AI assistant that makes phone calls
Fine-tune 20 Llama Models in 5 Minutes
The GenAI Maturity Curve, or You Probably Don't Need Fine-Tuning
Unveiling the Latest Gemma Model Advancements
Agentic Workflows on Vertex AI
Building security around ML
GitHub Next Explorations
GitHub's AI Powered Security Platform
Insights from Snorkel AI Running Azure AI Infrastructure
LLM Quality Optimization Bootcamp
RAG and the MongoDB Document Model
GitHub Copilot: The World's Most Widely Adopted AI Developer Tool
Llama 3 at 1,000 tok/s on the SambaNova AI Platform
Accelerate your AI journey with Azure AI model catalog
AI Templates
Build, Evaluate and Deploy a RAG-Based Retail Copilot with Azure AI
Creating and scaling your own custom copilots with Azure AI Studio
GitHub Copilot: The World's Most Widely Adopted AI Developer Tool
How to Add Secure Code Interpreting in Your AI App
Ionic Launch: Opening the economy to AI agents
Lessons from the Trenches: Building LLM Evals That Work IRL
Multi model multimodal and multi agent innovations in Azure AI
Which Jobs Can Be Replaced Today
Your RAG is Tripping, Here's the Real Reason Why
BotDojo Launch: Enhancing AI Assistants with Evaluations and Synthetic Data
Cohere for VPs of AI
Scaling AI in Education: A Khanmigo Case Study
Mastering LLM Inference Optimization From Theory to Cost Effective Deployment
AI Platform Engineering
Cooking with fire without burning down the kitchen
E-Values: Evaluating the Values of AI
Hiring & Building an AI Engineering Team
LLM Safeguards: Security, Privacy, Compliance, and Anti-Hallucination
Navigating Challenges and Technical Debt in LLMs Deployment
RAG for VPs of AI
Real ROI: Lessons from Enterprises That Have Already Succeeded with LLMs at Scale
The ROI of AI: Why You Need an Eval Framework
Understanding AI Stakes to Break Production Code
AI Frontiers in Trust and Safety: Combatting Multifaceted Harm on Tinder at Scale
Enhancing Quality and Security in CI
Iterating on LLM apps at scale: Learnings from Discord
Decoding Mistral AI's Large Language Models
The AI Emperor Has No DAUs: Why Most Developers Still Do Not Use Code AI
A Practical Guide to Efficient AI
Moondream: How Does a Tiny Vision Model Slap So Hard?
Navigating RAG Optimization with an Evaluation Driven Compass
How Zapier Builds AI Products and Features with the Help of Braintrust
What It Actually Takes to Deploy GenAI Applications to Enterprises
Knowledge Graphs & GraphRAG: Techniques for Building Effective GenAI Applications
AI Engineering Without Borders
State Space Models for Realtime Multimodal Intelligence
Second Order Effects of AI
Build an AI Research Agent
The Multimodal Future of Education
Productionizing GenAI Models
Code Generation and Maintenance at Scale
The Hierarchy of Needs for Training Dataset Development
No More Bad Outputs with Structured Generation
Architecting and Testing Controllable Agents
Realtime Data Connectivity for AI
Breaking AI's 1-GHz Barrier
Making Open Models 10x Faster and Better for Modern Application Innovation
Building and Scaling an AI Agent Swarm of Low-Latency Real-Time Voice Bots
We accidentally made an AI platform
Everything You Need to Know About Fine-tuning and Merging LLMs
The Era of Unbounded Products: Designing for Multimodal IO
LLM Scientific Reasoning: How to Make AI Capable of Nobel Prize Discoveries
How to Construct Domain Specific LLM Evaluation Systems
From Model Weights to API Endpoint with TensorRT-LLM
Build enterprise generative AI apps using Llama 3 at 1,000 tokens/s on the SambaNova AI platform
Going beyond RAG: Extended Mind Transformers
Judging LLMs
Pydantic is STILL all you need
More Nodes Is All You Need
GraphRAG: The Marriage of Knowledge Graphs and RAG
Building State of the Art Open Weights Tool Use: The Command R Family
Git push get an AI API
Hypermode Launch
Disrupting the $15 Trillion Construction Industry with Autonomous Agents
10x Development: LLMs for the Working Programmer
Building Reliable Agentic Systems
Building with Anthropic Claude: Prompt Workshop
Running AI Applications in Minutes with AI Templates
Decoding the Decoder LLM without de code
Using agents to build an agent company
What's new from Anthropic and what's next
Emergence Launch: AI Agents and the Future Enterprise
Fixing bugs in Gemma, Llama, and Phi 3
How Codeium Breaks Through the Ceiling for Retrieval
Low Level Technicals of LLMs
Copilots Everywhere
Unlocking Developer Productivity across CPU and GPU with MAX
From Software Developer to AI Engineer
Lessons From A Year Building With LLMs
Open Challenges for AI Engineering
Llamafile: Bringing AI to the Masses with Fast CPU Inference
The Future of Knowledge Assistants
The Making of Devin
From Text to Vision to Voice: Exploring Multimodality with OpenAI
Keynotes & Multimodality Track
GPUs & Inference Track
Keynotes & CodeGen Track
Open Models track
Open Models track