Back to latest

Tech Videos

Sources AI Engineer All-In Podcast Andrej Karpathy Anthropic Apple Apple Developer AWS Events ByteByteGo Computerphile Cursor Dwarkesh Patel EO Fireship GitHub Google Cloud …

Sources

Watch First

Deep dive on LLM Inference at Scale — Harshul Jain, Audible & Tanmay Sah, Independent AI Researcher by AI Engineer is hands-down the single video most worth your time today. It delivers a rigorous, first-principles engineering masterclass on LLM inference mechanics, detailing memory versus compute bounds, High Bandwidth Memory (HBM) transfer bottlenecks, and live H100 benchmarks comparing vLLM and SGLang.

Highlights by Theme

Developer Tools & Platforms

GitHub’s How to teach GitHub Copilot about your codebase | Tutorial for Beginners demonstrates repo-level context engineering using .github/copilot-instructions.md, file-type instruction scopes, agent skills, and Playwright MCP servers for browser automation. On Lex Clips, How agents are perfect for Linux | DHH and Lex Fridman illustrates how LLM agents systematically diagnose esoteric system errors by cross-referencing logs with application source code down to exact Rust file line numbers. Meanwhile, Syntax’s Openai Releases GPT 6 Astra ⟡ Vitest 5 is Here⟡ Audacity 4 Gets a facelift ⌁ Syntax Weekly ⌁ reviews Vitest 5 benchmark gains ranging from 8% to 53% alongside DaVinci Resolve’s new native Model Context Protocol (MCP) integration.

AI & Machine Learning

Leading with technical depth, AI Engineer’s Deep dive on LLM Inference at Scale — Harshul Jain, Audible & Tanmay Sah, Independent AI Researcher unpacks how Multi-Head Latent Attention (MLA) achieves a 14x memory compression over standard Multi-Head Attention while evaluating prefill compute bounds versus decode memory bounds on H100 GPUs. Google DeepMind’s AlphaGenome Atlas: Understanding the human genome highlights a 1-petabyte platform that precomputes Variant Impact Scores across all 9 billion single-letter genomic mutations. Additionally, AWS Developers’ AWS Partners with Anthropic on the Model Hardware Standard introduces the Model Hardware Standard (MHS) spec, extending tool calling like MCP to physical robotics through continuous sensory streaming and bidirectional control loops.

Hardware & Infrastructure

Bloomberg Tech’s Amazon, Qualcomm Deal Broadens the AI Chip Race | Bloomberg Tech 9/08/2026 explores Qualcomm’s multi-generation data center supply deal with Amazon—providing custom connectivity IP and AI inference chips—to diversify hyperscaler hardware away from Nvidia. The same coverage highlights TSMC and Samsung committing to ASML’s High-NA EUV lithography systems and 12-inch photomasks to scale advanced node manufacturing. Meanwhile, NVIDIA’s Building Europe’s AI Stack — Inside the Barcelona Supercomputing Center showcases MareNostrum 5 and Europe’s “AI Factory” blueprint, leveraging a five-layer stack across energy, chips, infrastructure, models, and enterprise applications.

Everything Else

On Lex Clips, DHH on fatherhood: What I love about being a father | Lex Fridman Podcast Clips steps away from software to discuss the profound personal fulfillment and perspective gained through parenting. Furthermore, Will AI kill most jobs? | DHH and Lex Fridman re-examines David Graeber’s “bullshit jobs” concept, suggesting recent tech sector downsizing reflects post-pandemic normalization rather than immediate AI job displacement.


💡 Would you like me to create a tailored summary or slide deck covering any specific technical topic from today’s videos, such as LLM inference optimizations or robotics control protocols?

Search MacWorks

Enter at least two characters.