YOUTUBE
Tech Videos
Sources AI Engineer All-In Podcast Andrej Karpathy Anthropic Apple Apple Developer AWS Events ByteByteGo Computerphile Cursor Dwarkesh Patel EO Fireship GitHub Google Cloud …
Sources
- AI Engineer
- All-In Podcast
- Andrej Karpathy
- Anthropic
- Apple
- Apple Developer
- AWS Events
- ByteByteGo
- Computerphile
- Cursor
- Dwarkesh Patel
- EO
- Fireship
- GitHub
- Google Cloud Tech
- Google DeepMind
- Google for Developers
- Hung-yi Lee
- Lenny's Podcast
- Lex Clips
- Lex Fridman
- Life at Google
- Marques Brownlee
- Microsoft
- No Priors: AI, Machine Learning, Tech, & Startups
- Numberphile
- NVIDIA
- OpenAI
- Perplexity
- Quanta Magazine
- Slack
- The Pragmatic Engineer
- Visual Studio Code
Watch First
Deep dive on LLM Inference at Scale — Harshul Jain, Audible & Tanmay Sah, Independent AI Researcher by AI Engineer is hands-down the single video most worth your time today. It delivers a rigorous, first-principles engineering masterclass on LLM inference mechanics, detailing memory versus compute bounds, High Bandwidth Memory (HBM) transfer bottlenecks, and live H100 benchmarks comparing vLLM and SGLang.
Highlights by Theme
Developer Tools & Platforms
GitHub’s How to teach GitHub Copilot about your codebase | Tutorial for Beginners demonstrates repo-level context engineering using .github/copilot-instructions.md, file-type instruction scopes, agent skills, and Playwright MCP servers for browser automation. On Lex Clips, How agents are perfect for Linux | DHH and Lex Fridman illustrates how LLM agents systematically diagnose esoteric system errors by cross-referencing logs with application source code down to exact Rust file line numbers. Meanwhile, Syntax’s Openai Releases GPT 6 Astra ⟡ Vitest 5 is Here⟡ Audacity 4 Gets a facelift ⌁ Syntax Weekly ⌁ reviews Vitest 5 benchmark gains ranging from 8% to 53% alongside DaVinci Resolve’s new native Model Context Protocol (MCP) integration.
AI & Machine Learning
Leading with technical depth, AI Engineer’s Deep dive on LLM Inference at Scale — Harshul Jain, Audible & Tanmay Sah, Independent AI Researcher unpacks how Multi-Head Latent Attention (MLA) achieves a 14x memory compression over standard Multi-Head Attention while evaluating prefill compute bounds versus decode memory bounds on H100 GPUs. Google DeepMind’s AlphaGenome Atlas: Understanding the human genome highlights a 1-petabyte platform that precomputes Variant Impact Scores across all 9 billion single-letter genomic mutations. Additionally, AWS Developers’ AWS Partners with Anthropic on the Model Hardware Standard introduces the Model Hardware Standard (MHS) spec, extending tool calling like MCP to physical robotics through continuous sensory streaming and bidirectional control loops.
Hardware & Infrastructure
Bloomberg Tech’s Amazon, Qualcomm Deal Broadens the AI Chip Race | Bloomberg Tech 9/08/2026 explores Qualcomm’s multi-generation data center supply deal with Amazon—providing custom connectivity IP and AI inference chips—to diversify hyperscaler hardware away from Nvidia. The same coverage highlights TSMC and Samsung committing to ASML’s High-NA EUV lithography systems and 12-inch photomasks to scale advanced node manufacturing. Meanwhile, NVIDIA’s Building Europe’s AI Stack — Inside the Barcelona Supercomputing Center showcases MareNostrum 5 and Europe’s “AI Factory” blueprint, leveraging a five-layer stack across energy, chips, infrastructure, models, and enterprise applications.
Everything Else
On Lex Clips, DHH on fatherhood: What I love about being a father | Lex Fridman Podcast Clips steps away from software to discuss the profound personal fulfillment and perspective gained through parenting. Furthermore, Will AI kill most jobs? | DHH and Lex Fridman re-examines David Graeber’s “bullshit jobs” concept, suggesting recent tech sector downsizing reflects post-pandemic normalization rather than immediate AI job displacement.
💡 Would you like me to create a tailored summary or slide deck covering any specific technical topic from today’s videos, such as LLM inference optimizations or robotics control protocols?