A curated list of AI-native products and resources, where an LLM is the product, not a feature.
The bar is deliberately high. "AI-powered" sidebars, summarize buttons, and semantic search bolted onto legacy products do not belong here. Links point at the product itself, and entries are dropped once a product shuts down or pivots away from AI.
- Chat & agents
- Vibe-coding and vibe-engineering
- Knowledge work
- Industry applications
- Creative & media
- Models & robotics
- Compute & infrastructure
- Developer tooling
- Learning & community
Consumer assistants and autonomous agents you interact with directly. These are the products most people picture when they say "AI," from plain chat interfaces to agents that browse, act, and drive a computer on your behalf.
General-purpose conversational assistants you talk to across web, mobile, and desktop.
- ChatGPT - OpenAI's consumer chat product with memory, canvas, voice, and the broadest third-party app ecosystem.
- Claude - Anthropic's assistant known for long-context reasoning, Projects, Artifacts, and Computer Use.
- Gemini - Google's multimodal assistant integrated across Workspace, Android, and the Pixel line.
- Grok - xAI's assistant with real-time access to X and a distinct, less-filtered persona.
- Meta AI - Assistant embedded across WhatsApp, Instagram, Messenger, and Ray-Ban Meta glasses.
- Vibe - Mistral's assistant, formerly Le Chat, unifying chat, work automation, and remote coding agents.
Agents that plan and execute multi-step tasks across your tools, often specialized to a company's workflows and grounded in private data, returning finished work rather than just answers.
- ChatGPT Agent - OpenAI's agent that browses, uses tools, and completes multi-step web tasks on the user's behalf.
- Claude Cowork - Anthropic's agent for non-coding knowledge work that runs background and scheduled tasks across shared projects and files.
- Dust - Enterprise platform for building no-code AI agents grounded in company data through 100+ connectors, where teams and agents collaborate in a shared workspace.
- Manus - General-purpose agent that plans, browses, writes code, and returns finished deliverables.
- Swiftask - No-code platform for building and orchestrating multi-model AI agents across a company's tools, with centralized governance.
Open-source assistants you self-host on your own server or VPS rather than your laptop, with persistent memory, custom skills, and access over the messaging apps you already use.
- Hermes Agent - Nous Research's self-hosted agent that keeps memory across sessions, writes its own reusable skills, and is reachable over Telegram, Discord, Slack, WhatsApp, Signal, and CLI.
- OpenClaw - Self-hosted personal AI assistant that runs on your own devices and answers you across WhatsApp, Telegram, Slack, Discord, Signal, iMessage, and many other channels.
- Paperclip - Open-source, self-hosted control plane for running and governing teams of AI agents with org charts, budgets, and goals.
Models and agents that operate a real browser or desktop by reading the screen and taking actions the way a person would.
- Browser Use - Open-source library that lets AI agents control a browser to complete web tasks.
- Claude for Chrome - Anthropic's Claude agent that navigates and takes actions inside your Chrome browser.
- cua - Open-source infrastructure and SDK for building computer-use agents that control full desktops in sandboxes.
- Firecrawl - API that crawls and scrapes websites into clean, LLM-ready markdown and structured data.
- H Company - Lab whose Holo models and Surfer H agent operate a real browser and desktop by reasoning, planning, and executing multi-step computer-use tasks.
- OpenAI Operator - OpenAI's computer-use agent that operates a browser via screenshots and actions.
Tools that let you build software by describing intent rather than writing every line. This covers coding agents, repo-understanding assistants, and prompt-to-app builders.
Agents and AI-first editors that read, write, and run code against real repositories.
- Claude Code - Anthropic's agentic tool that reads, edits, and executes against a real codebase from the terminal or IDE.
- Cline - Open-source autonomous coding agent that runs inside VS Code with human-in-the-loop approval.
- Cursor - AI-first code editor built around whole-repo context, an agent mode, and multi-line tab completion.
- Devin - Cognition's autonomous software engineer that takes tasks from Slack or an issue tracker, then plans, codes, tests, and opens pull requests.
- Gemini CLI - Google's open-source terminal agent powered by Gemini, with MCP support and a generous free tier.
- GitHub Copilot - Inline completion and an agent mode that edits across files and opens pull requests, integrated across GitHub and popular IDEs.
- Google Antigravity - Agent-first development platform built around Gemini that orchestrates autonomous agents across editor, terminal, and browser.
- Mistral Vibe - Mistral's terminal and IDE coding agent with cloud-based remote agents, built on the open-source mistral-vibe CLI.
- OpenAI Codex - OpenAI's coding agent spanning CLI, IDE, cloud, and web.
- OpenCode - Provider-agnostic, open-source terminal coding agent.
- Poolside - Enterprise coding assistant powered by its own foundation models, deployed inside a customer's environment and tuned on their codebases.
- Windsurf - Agentic IDE with Cascade, a multi-step coding agent that operates across files.
- Zed - Collaborative editor written in Rust with first-class agent panels and multi-model support.
Tools that turn a codebase into searchable docs, wikis, and answers for humans and agents alike.
- Code Wiki - Google's tool that turns any GitHub repo into an interactive, always-in-sync wiki with architecture diagrams and a Gemini-powered chat.
- Context7 - Upstash service that feeds up-to-date, version-specific library documentation into coding agents over MCP.
- DeepWiki - Cognition's Devin-generated wikis for any public GitHub repo, with architecture diagrams, source links, and a conversational assistant.
- Zread - Generates structured guides, API docs, and code Q&A for any GitHub repo by swapping github.com for zread.ai.
Pure vibe-coding platforms that turn a plain-language description into a working app, aimed at non-technical builders who never touch the underlying code.
- Bolt - Prompt-to-app builder that generates and runs full web apps directly in the browser.
- Figma Make - Prompt-to-app tool inside Figma that generates working prototypes from a description.
- Google AI Studio - Browser-based platform for prototyping and building apps with Gemini models across text, image, audio, and video.
- Google Opal - Google Labs no-code builder for creating and sharing mini AI apps as visual node workflows.
- Google Stitch - Gemini-powered tool that turns prompts or wireframes into high-fidelity UI plus front-end code, absorbing the former Galileo AI.
- Lovable - Prompt-to-app builder that ships full-stack web apps from a chat window.
- Replit Agent - Browser-based agent that builds, runs, and deploys full-stack applications from natural-language prompts.
- v0 - Vercel's generative UI tool that produces production-ready React and Tailwind components.
AI-native products that help you find, organize, and act on information. Search and research engines, note-taking systems, and meeting assistants live here.
Answer engines and retrieval APIs that find, cite, and synthesize sources instead of returning a list of links.
- Claude Science - Anthropic research workbench for scientists, pre-configured with scientific databases, coding tools, and compute.
- Exa - Neural search API designed for agents and retrieval-heavy applications rather than human browsing.
- Perplexity - Answer engine that retrieves, cites, and synthesizes web sources in place of a traditional results page.
- You.com - Multi-mode AI search with agent workflows and enterprise deployments.
Notes and knowledge bases that ground AI in your own documents and surface what's relevant.
- Gemini Notebook (formerly NotebookLM) - Google's research and note-taking tool that grounds answers in your uploaded sources and generates audio and video overviews.
- Mem - Self-organizing notes app that surfaces related notes and drafts follow-ups on demand.
- Notion MCP - Notion's official MCP server that lets AI agents and assistants search, read, and update a Notion workspace.
- Obsidian - Local-first markdown knowledge base with a large plugin ecosystem for AI assistants, semantic search, and note generation.
Assistants that record, transcribe, and summarize meetings, then draft the follow-ups.
- Fathom - Meeting recorder that produces action items, CRM updates, and follow-up emails from a conversation.
- Fireflies - Meeting assistant that records, transcribes, and answers questions across a team's call history.
- Granola - Notepad that listens to your meetings and turns your rough notes into structured summaries.
- Otter - Live transcription and meeting summaries with a chat interface over meeting history.
Vertical products where an LLM sits at the core of a specific profession's workflow, from customer support to law, medicine, sales, and the classroom.
AI agents that resolve customer conversations end to end across chat and messaging channels.
- Hugo - Crisp's AI support agent that resolves conversations end-to-end across chat and messaging channels, connecting to live data and actions through MCP.
- Intercom Fin - AI support agent built on Intercom's messenger that answers questions and takes actions across connected systems.
AI-native CRMs and outbound platforms that research prospects, enrich records, and run sequences.
- 11x - Digital sales development reps that prospect, personalize, and book meetings.
- Artisan - Outbound platform fronted by Ava, an AI SDR that researches and emails prospects.
- Attio - AI-native CRM that builds pipeline views, enriches records, and runs research agents across customer data.
- Clay - Data enrichment and outbound workspace that composes dozens of providers and LLM prompts into a spreadsheet.
- muchbetter.ai - AI sales-coaching platform that trains field teams through voice-driven role-play simulations with virtual clients and real-time feedback.
- Unify - Warm outbound platform that combines intent signals with AI-written sequences.
Platforms for legal research, drafting, review, and litigation analytics used inside law firms.
- Harvey - Legal workflow platform deployed across large law firms for drafting, diligence, and case analysis.
- Ironclad - Contract review and redlining built into Ironclad's CLM.
- Legora - Collaborative legal AI platform for research, drafting, and review, adopted across law firms worldwide.
- Lex Machina - Legal analytics platform mining litigation data to predict case outcomes and inform strategy.
Clinical documentation, medical answer engines, and AI applied to drug discovery and diagnostics.
- Abridge - Ambient clinical documentation used in major health systems to generate notes during patient visits.
- Isomorphic Labs - Alphabet/DeepMind spinout applying AI foundation models to drug discovery and molecular design.
- Nabla - Ambient AI assistant for clinicians that writes structured notes from the consultation audio.
- OpenEvidence - Medical answer engine for clinicians grounded in peer-reviewed literature.
- Owkin - AI biotech for drug discovery and diagnostics, training multimodal models on data from hospitals worldwide.
- Suki - Voice-first clinical assistant that drafts notes, orders, and codes.
Tutors and teaching assistants that guide learners with Socratic prompting and help educators plan and grade.
- ChatGPT Study Mode - ChatGPT mode that guides learners through problems step by step with Socratic prompts instead of giving direct answers.
- Claude Learning Mode - Claude style that guides you to your own answer with a Socratic approach instead of solving directly, with Explanatory and Learning variants in Claude Code.
- Duolingo Max - Duolingo's premium tier adding AI-powered roleplay conversations and answer explanations for language learning.
- Khanmigo - Khan Academy's AI tutor that guides students through problems using Socratic prompting.
- MagicSchool - AI platform for teachers that plans lessons, differentiates materials, and drafts feedback.
- Speak - AI-powered language tutor built around spoken conversation practice with real-time feedback.
Generative tools for producing and editing visual, audio, and video content, from design canvases to image, video, and music models.
Generative design canvases that turn prompts into editable visuals, prototypes, and production assets.
- Claude Design - Anthropic Labs tool that turns prompts into designs, prototypes, slides, and decks and can apply your team's design system.
- Krea - Real-time generative canvas for image, video, and 3D creation.
- Pencil - AI-native design canvas for developers that generates editable designs and production-ready code from a VS Code or Cursor integration.
- Photoroom - AI photo editor and product-image studio with instant background removal and generative editing.
- Weavy - Node-based creative canvas that chains AI models for image, video, and 3D generation and editing.
Text-to-image models and editors for creating and refining still images.
- Flux - Black Forest Labs' open-weight and hosted models known for photoreal prompt adherence.
- Imagen - Google DeepMind's text-to-image model generating high-resolution images with SynthID watermarking.
- Magnific - AI upscaler and enhancer that adds detail and reimagines images at high resolution.
- Midjourney - Generative image service known for a consistent aesthetic and a large creator community.
- Nano Banana - Google's Gemini-based image model line known for precise, conversational editing.
- Seelab - AI photo studio that trains custom models on your products and brand to generate on-brand product visuals.
Models and studios that generate and edit video, some with natively synchronized audio.
- Google Flow - AI filmmaking studio built on Veo and Imagen for generating and editing cinematic scenes.
- HeyGen - Generates talking-avatar videos from a script, with voice cloning and translation for marketing and training.
- Kling - Kuaishou's video model with long-duration generations and strong motion coherence.
- Pika - Generative video app with effects-driven editing and social sharing.
- Runway - Video model suite and editor used in film and advertising production.
- Veo - Google DeepMind's video model producing high-resolution clips with natively generated synchronized audio.
Speech-to-text, text-to-speech, voice cloning, and full music generation.
- AssemblyAI - Speech-to-text and audio-intelligence API with transcription, diarization, and LLM-powered audio understanding.
- Deepgram - Voice AI API for fast, accurate speech-to-text and text-to-speech built for real-time agents.
- ElevenLabs - Voice synthesis platform covering cloning, dubbing, conversational agents, and audiobooks.
- Gladia - Audio infrastructure API for real-time and async speech-to-text, diarization, and translation across 100+ languages.
- Gradium - Kyutai spinout building ultra-low-latency text-to-speech, speech-to-text, and voice-cloning models for voice agents, behind a single API.
- Suno - Text-to-song model that generates full vocal tracks from a prompt.
- Udio - Music generation model with fine control over style and structure.
The models themselves and the machines they drive: open-weight and small language models, OCR, and robotics foundation models for embodied AI.
LLM families released with downloadable weights you can run, fine-tune, and self-host.
- DeepSeek - Lab releasing reasoning- and coding-focused models with weights published under the MIT License.
- Gemma - Google DeepMind's family of lightweight open models built from the same research as Gemini, released under a permissive license.
- GLM - Open-weight General Language Model family from Z.ai, formerly Zhipu AI, targeting long-horizon agentic and coding tasks.
- gpt-oss - OpenAI's open-weight reasoning models released under Apache 2.0 and designed to run on single-GPU or consumer hardware.
- Kimi - Moonshot AI's Kimi family of large mixture-of-experts open-weight models, strong on agentic and coding tasks.
- Llama - Meta's widely deployed open model family, with natively multimodal mixture-of-experts architectures.
- MiniMax - Lab releasing frontier-class open-weight reasoning and coding models under MIT-style licenses.
- Mistral - Lab releasing open-weight dense and mixture-of-experts models under Apache 2.0.
- Nemotron - NVIDIA's open model family released with open weights, training data, and recipes for agentic reasoning.
- Qwen - Alibaba's open-weight family, one of the most-downloaded open model lines, spanning phone-sized to very large MoE models.
- Whisper - OpenAI's open-source automatic speech recognition model for multilingual transcription and translation.
Compact models tuned to run efficiently on-device or on modest hardware.
- AlphaEdge - Sovereign-AI startup that compresses and distills models into small, efficient ones for on-prem document and enterprise workflows.
- Phi - Microsoft's family of small open-weight models emphasizing strong reasoning at small parameter counts.
- Pleias - Lab training small open-weight models exclusively on public-domain and permissibly licensed data, alongside the fully open Common Corpus dataset.
- SmolLM - Hugging Face's fully open family of small text and vision models, with training data and code released for on-device use.
Foundation models and companies bringing learned control to humanoids and real-world machines.
- 1X - Humanoid robotics company building the Neo home robot around neural control policies.
- AMI Labs - Yann LeCun's lab building world models that learn abstract representations of the physical world so agents can predict and plan their actions.
- Figure - Humanoid robotics company training general-purpose manipulation models for commercial deployment.
- Genesis AI - Physical-AI lab building robotics foundation models and a general-purpose robot that reasons, plans, and acts in the real world.
- Gobano Robotics - Robotics company applying imitation learning, reinforcement learning, and world models to dexterous industrial and logistics tasks like folding, assembling, and bin picking.
- Physical Intelligence - Research lab building foundation models for robot control across embodiments.
- Skild AI - Developer of a general-purpose robot brain trained across thousands of hours of robot data.
- UMA - Physical-AI company building humanoid and mobile robots that learn new skills from demonstration.
- Wandercraft - Robotics company building AI-powered self-balancing exoskeletons and the Calvin humanoid robot for industrial tasks.
Models that convert documents and images into structured, machine-readable text.
- DeepSeek-OCR - Open-weight OCR model using optical token compression for high-throughput, low-cost document-to-Markdown conversion; runs on vLLM.
- Mistral OCR - Hosted OCR API with strong accuracy on complex tables and handwriting for managed, no-ops document extraction.
The hardware and platforms that train and serve models: AI chips, local inference engines, GPU clouds, and sandboxes for running agent-generated code.
Processors purpose-built for training and inference, from GPUs to wafer-scale and photonic chips.
- Arago - Deeptech building a photonic AI chip that runs matrix math with light for order-of-magnitude lower energy use.
- Cerebras - Builder of wafer-scale chips that hold an entire model on a single processor for ultra-fast training and inference.
- Groq - Designer of the LPU inference chip and cloud known for sub-hundred-millisecond token generation.
- NVIDIA - Dominant maker of AI GPUs and the CUDA software stack that most frontier training and inference runs on.
- Sesterce - AI-factory company operating large, sustainable GPU clusters for training and inference.
- VSORA - Startup building an inference chip to challenge GPU makers on performance per watt.
Engines and apps for running models on your own CPUs and GPUs.
- llama.cpp - LLM inference engine in pure C/C++ that runs GGUF-format models across CPUs and GPUs on a wide range of hardware.
- LM Studio - Desktop app to browse, download, and run local models with a built-in chat interface and local server.
- MLX - Apple's array and machine-learning framework optimized for Apple silicon, for on-device training and inference.
- Ollama - Tool that packages and runs LLMs locally via a simple command line and local API.
- vLLM - High-throughput inference and serving library using PagedAttention and continuous batching for production deployment.
- ZML - Zig-based inference stack that compiles models into standalone native binaries across NVIDIA, AMD, TPU, and Trainium with zero Python dependencies.
GPU clouds, hosted inference, and fine-tuning platforms for deploying models at scale.
- Adaptive ML - Reinforcement-learning platform for post-training, evaluating, and serving enterprise-specialized open models.
- CoreWeave - GPU cloud purpose-built for AI, offering large-scale NVIDIA clusters for training and inference.
- Fireworks AI - Inference and fine-tuning platform for open-weight models with low-latency routing.
- FlexAI - Universal AI compute platform that abstracts heterogeneous hardware so developers can train, fine-tune, and serve models without managing infrastructure.
- Hugging Face - The hub for open models, datasets, and Spaces, plus the Transformers library and hosted inference that anchor the open-source AI ecosystem.
- Koyeb - Serverless platform for deploying AI inference and apps on autoscaling CPUs and GPUs worldwide.
- Lambda - GPU cloud offering on-demand and reserved NVIDIA clusters for training and inference.
- llm-d - Red Hat-led, Kubernetes-native distributed inference framework built on vLLM for serving LLMs at scale.
- Nebius - AI-focused cloud providing GPU compute, managed inference, and ML tooling.
- NVIDIA Dynamo - Open-source datacenter-scale orchestration layer that coordinates vLLM, SGLang, and TensorRT-LLM into distributed multi-node inference.
- NVIDIA NIM - Prebuilt, containerized inference microservices for deploying optimized model endpoints across cloud, data center, and workstation.
- OpenRouter - Unified API and marketplace routing requests across hundreds of models and providers.
- Replicate - Hosted registry and inference layer for open-source models.
- RunPod - GPU cloud with on-demand and serverless GPUs for training and inference.
- Together AI - Inference, fine-tuning, and training platform for open-weight models.
Isolated, ephemeral environments for safely executing AI-generated code.
- Daytona - Secure, elastic infrastructure for running AI-generated code, with agent sandboxes that spin up in tens of milliseconds.
- E2B - Open-source runtime that gives AI agents isolated cloud sandboxes to execute code.
- Modal - Serverless Python platform offering on-demand sandboxes for running AI-generated code, model inference, and batch jobs.
- Vercel Sandbox - Ephemeral, isolated microVMs for running untrusted AI-generated code, from Vercel.
Building blocks for developers shipping AI features: agent frameworks, MCP integrations, vector search, memory layers, and evaluation and observability.
Libraries for building, orchestrating, and observing LLM apps and multi-agent systems.
- Claude Agent SDK - Anthropic's SDK for building custom agents on the same harness that powers Claude Code, with tools, subagents, and MCP support; formerly the Claude Code SDK.
- CrewAI - Open-source Python framework for orchestrating role-playing, collaborative multi-agent systems.
- DeepAgents - LangChain's batteries-included agent harness with planning, sub-agents, a virtual filesystem, and context management for long-running tasks.
- Google ADK - Google's open-source Agent Development Kit for building, evaluating, and deploying multi-agent systems, model-agnostic and optimized for Gemini.
- LangChain - Framework and platform for building, running, and observing LLM applications.
- LlamaIndex - Data framework for connecting LLMs to private and enterprise data.
- Vercel AI SDK - TypeScript toolkit for building streaming AI applications across providers.
Platforms for building and connecting Model Context Protocol servers and agent tool integrations.
- Alpic - MCP-native cloud platform to build, deploy, monitor, and distribute MCP servers and ChatGPT apps.
- Composio - Integration layer that connects AI agents to over a thousand tools with managed authentication, tool search, and agent-friendly APIs.
Embedding stores and search engines that power retrieval-augmented generation.
- Chroma - Open-source embedding database focused on developer ergonomics.
- LanceDB - Embedded and serverless vector database built on the Lance columnar format.
- Meilisearch - Open-source search engine with hybrid full-text and vector search and built-in AI-powered semantic ranking.
- Pinecone - Managed vector database used in production retrieval pipelines.
- Qdrant - Open-source vector search engine with on-disk indexing and filtering.
- Turbopuffer - Serverless vector and full-text search built directly on object storage.
- Weaviate - Open-source vector database with hybrid search and built-in modules.
Memory layers that give agents persistent, personalized context across sessions.
- Cognee - Open-source memory layer that builds knowledge graphs from your data for AI agents and RAG.
- Letta - Platform for stateful agents with long-term memory, formerly the MemGPT project.
- Mem0 - Open-source memory layer that gives AI agents persistent, personalized memory across sessions.
- Supermemory - Single memory API for fact extraction, user-profile building, contradiction resolution, and selective forgetting.
- Zep - Memory server for AI agents built on a temporal knowledge graph via the open-source Graphiti library.
Tracing, testing, and monitoring to measure and debug LLM and agent behavior in production.
- Datadog LLM Observability - End-to-end tracing, monitoring, and quality and safety evaluation for LLM and agentic applications.
- Giskard - Open-source testing and evaluation framework that detects hallucinations, biases, and vulnerabilities in LLM and ML systems.
- Helicone - Open-source LLM observability proxy and analytics.
- Langfuse - Open-source tracing, evals, and prompt management for LLM apps.
- LangSmith - Tracing, evaluation, and monitoring for LangChain and plain LLM apps.
- Ragas - Open-source evaluation framework for RAG and LLM applications, with metrics for faithfulness, answer relevance, and context precision.
Where to keep up and go deeper: newsletters, podcasts, people to follow, courses, foundational papers, events, and communities.
Writers and publications covering AI engineering, research, and industry economics.
- Chip Huyen - Essays on ML systems, AI engineering, and shipping models to production.
- Eugene Yan - Applied scientist writing on ML systems design, LLM evals, and recommender systems.
- Hamel Husain's Blog - Practical writing on LLM evaluation, fine-tuning, and shipping ML products.
- Han-Chung Lee - Notes on ML engineering, evaluation, and compound AI systems.
- Latent Space - Practitioner-focused newsletter and community covering AI engineering.
- Lilian Weng - In-depth technical posts on agents, RL, diffusion, and LLMs.
- One Useful Thing - Ethan Mollick's essays on how knowledge workers are actually using frontier models.
- Sebastian Raschka - "Ahead of AI": deep dives on LLM training, architectures, and research.
- SemiAnalysis - Dylan Patel's research on AI hardware, data centers, and the semiconductor supply chain behind frontier compute.
- Simon Willison's Weblog - Detailed technical notes on LLMs, tools, and prompt engineering from the creator of Datasette.
- Tensor Economics - Analysis of the economics of AI compute, inference costs, and model serving.
- Where's Your Ed At - Ed Zitron's newsletter with skeptical, deeply reported takes on AI hype, economics, and Big Tech.
Long-form conversations with the researchers and founders building frontier AI.
- Dwarkesh Podcast - Long-form interviews with researchers and founders on where AI is heading.
- Latent Space Podcast - Conversations with AI engineers on the systems, tools, and tradeoffs behind production apps.
- No Priors - Elad Gil and Sarah Guo interview the builders of frontier AI companies.
- The Cognitive Revolution - Deep technical interviews on frontier model capability and alignment.
Researchers, builders, and educators worth following for signal on where AI is going.
- Andrej Karpathy - Founding OpenAI member and former Tesla AI lead whose talks and threads explain LLMs from first principles.
- Andrew Ng - DeepLearning.AI and Coursera founder posting pragmatically on applied AI, agents, and education.
- Boris Cherny - Creator of Claude Code at Anthropic, sharing tips and thinking on agentic coding.
- Chip Huyen - Author of "AI Engineering" and "Designing Machine Learning Systems" writing on shipping ML and LLM systems to production.
- Clément Delangue - Co-founder and CEO of Hugging Face, championing open-source AI and the model-sharing community.
- François Chollet - Keras creator and ARC-AGI co-author on reasoning, generalization, and where deep learning falls short.
- Jim Fan - NVIDIA senior research lead covering embodied AI, robotics, and foundation agents.
- Lilian Weng - Former OpenAI research lead whose long-form posts are canonical references on agents, RL, and diffusion.
- Sebastian Raschka - Author of "Build a Large Language Model (From Scratch)" who breaks down model internals and training.
- Yann LeCun - Turing Award laureate and deep-learning pioneer arguing for world models and open research.
Hands-on courses and guides for learning to build and understand LLMs from the ground up.
- Claude Code Best Practices - Anthropic's engineering guide to agentic coding workflows and getting the most out of Claude Code.
- Claude Code Documentation - Official docs for Anthropic's terminal coding agent, covering setup, workflows, hooks, MCP, and the SDK.
- DeepLearning.AI Short Courses - Free hands-on courses on RAG, agents, evaluation, and multimodal systems.
- LLM101n - Andrej Karpathy's course building a Storyteller LLM from scratch, end to end.
- nanochat - Karpathy's minimal full-stack ChatGPT-style pipeline (tokenizer, pretraining, SFT, RL, inference) in one clean codebase.
- Neural Networks: Zero to Hero - Andrej Karpathy's video course that builds backprop, GPTs, and tokenizers from scratch.
- Stanford CS336 - Stanford's "Language Modeling from Scratch" course spanning tokenization, architecture, training, and systems.
- The Rise of the AI Engineer - The essay that named the discipline and remains a canonical reading list.
The research papers that define modern LLMs: architecture, scaling, reasoning, and post-training.
- Attention Is All You Need - Vaswani et al., 2017. The transformer: self-attention, multi-head attention, and positional encoding, the architecture everything else builds on.
- Chain-of-Thought Prompting - Wei et al., 2022. Showed reasoning can be elicited by prompting alone; the conceptual seed for later reasoning work.
- DeepSeek-R1 - Guo et al., 2025. Reasoning behavior emerging from reinforcement learning (GRPO) rather than prompting or SFT; the current post-training frontier.
- FlashAttention - Dao et al., 2022. IO-aware exact attention that treats the GPU memory hierarchy as the real bottleneck, which is why long context and cheap inference are practical.
- Kimi k1.5: Scaling RL with LLMs - Moonshot AI, 2025. The other major RL-reasoning report alongside DeepSeek-R1, with more long-context and infrastructure detail; together they define the reasoning-model recipe.
- RULER - Hsieh et al., 2024. Benchmark measuring the real usable context length of long-context models beyond simple retrieval.
- Training Compute-Optimal LLMs (Chinchilla) - Hoffmann et al., 2022. Showed token count matters as much as parameter count for a fixed compute budget, revealing most prior models were badly undertrained.
- Training LMs to Follow Instructions (InstructGPT) - Ouyang et al., 2022. The post-training blueprint, SFT then RLHF, that turned a text predictor into an assistant.
Conferences and hackathons for practitioners and researchers shipping AI.
- AI Engineer Summit - Annual conference for practitioners shipping AI into production.
- dotAI - AI conference for builders and developers, gathering full AI teams and top researchers.
- Gen AI Days - Conference on scaling generative AI systems in production, with real-world implementation talks.
- ICML - International Conference on Machine Learning, a top-tier ML research venue.
- NeurIPS - Conference on Neural Information Processing Systems, the largest ML research conference.
- NVIDIA GTC - NVIDIA's GPU Technology Conference, a major event for AI compute, research, and tooling.
- Shift Hackathon - 48-hour generative-AI hackathon.
Forums and subreddits where AI builders and enthusiasts share and discuss.
- r/artificial - General artificial-intelligence discussion community.
- r/ArtificialInteligence - Large general AI community (the subreddit's name is spelled this way).
- r/ChatGPT - One of the largest AI subreddits, covering ChatGPT and consumer AI.
- r/Claude - Community discussing Anthropic's Claude models and apps.
- r/ClaudeAI - Community for Claude and Claude Code users sharing workflows, prompts, and tips.
- r/codex - Community around OpenAI Codex and its coding workflows.
- r/cursor - Community for the Cursor AI code editor.
- r/LocalLLaMA - Running, fine-tuning, and evaluating open-weight models locally.
- r/MachineLearning - Long-running research-focused subreddit for ML papers and releases.
- r/mcp - Community around the Model Context Protocol and MCP servers.
- r/midjourney - Community for the Midjourney image generator.
- r/MistralAI - Community around Mistral's open-weight models and products.
- r/OpenAI - Community tracking OpenAI's models, products, and announcements.
- r/singularity - High-traffic subreddit debating AGI progress and frontier AI news.
- r/StableDiffusion - Hub for open image-generation models, workflows, and tooling.