All AI tools — index
Reference: every tool, side by side — look it up any time
1Overview
An honest survey of the AI tools your peers actually use — Claude Code, Perplexity, Gemini, Cursor, opencode, Lovable, v0, Base44. Switch between the matrix (task × tool fit), the grid (one tile per tool), and the day-by-day list. Click any tile to read a one-pager: what it does well, where it falls short, what to try in 5 minutes.
In this chapter you’ll explore a complete, side‑by‑side reference of the AI tools that professionals actually rely on—from Claude Code and Perplexity to Gemini and Cursor. You can flip between three views: a matrix that matches tasks to the best‑fit tool, a grid that shows each tool as an individual tile, and a chronological list that tracks daily usage. Clicking any tile opens a concise one‑page summary that tells you what the tool excels at, its limitations, and a quick five‑minute experiment you can try right away.
You can toggle between a matrix that maps tasks to tool fit, a grid showing one tile per tool, or a day‑by‑day list to see which fits your workflow best.
Clicking any tile opens a one‑pager that explains what the tool excels at, where it falls short, and suggests a concrete 5‑minute experiment you can try.
2Every tool
| I want to… | ▶ 1 | ▶ 3 | ▶ 3 | ▶ 3 | ▶ 2 | ▶ 3 | ▶ 2 | ▶ 2 | ▶ 2 | ▶ 9 | ▶ 3 | ▶ 2 | ▶ 8 | ▶ 2 | ▶ 3 | ▶ 4 | ▶ 3 | ▶ 3 | ▶ 3 | ▶ 2 | |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
Write code Python, R, JS — analyze, debug, visualize | |||||||||||||||||||||
Find papers Find scientific sources and cite them correctly | |||||||||||||||||||||
Draft docs Reports, slides, emails, lab notes — fast | |||||||||||||||||||||
Build an app Web tool or lab interface, no coding background needed | |||||||||||||||||||||
Automate workflow Clean data, rename files, script repetitive tasks | |||||||||||||||||||||
Talk through ideas Discuss papers, test ideas, get concepts explained | |||||||||||||||||||||
Open source / self-host | |||||||||||||||||||||
Data analytics | |||||||||||||||||||||
Templates | ●●● | ●● | ●●● | ●●● | ●●● | ||||||||||||||||
Learning resources |
In this course
Dify Open-source LLM app builder with RAG and agent orchestration
Groq Blazing-fast LLM inference — Llama and Gemma at 500+ tokens/sec
Cerebras The fastest LLM inference anywhere — open models at 1,800–3,000 tokens/sec
n8n Visual workflow automation — the lab bench for this course
OpenRouter One API, every major LLM — including many free models
LM Studio A polished desktop app to download and chat with local models
Chat & research
Claude Desktop All of Claude in a native app — works with your local files
Claude Cowork Delegate whole tasks to Claude — finished work, no terminal
Gemini Google's AI — embedded directly in Docs, Sheets, and Gmail
Perplexity AI search engine that cites real sources — not just summaries
Claude Claude in your browser — chat, write, analyze, and code
Jan An open-source, offline-first ChatGPT alternative for your desktop
Onyx Open-source AI chat + search over your team's apps
AnythingLLM A private, all-in-one ChatGPT for your own documents
📓 Open Notebook An open-source, privacy-focused alternative to Google's Notebook LM
ChatGPT OpenAI's assistant — broadest features, strong data analysis
💬 Open WebUI The ChatGPT-style front door your whole team logs into
🧮 Abacus.ai One bill, 25+ models, plus an agent that builds and one that lives in WhatsApp
🐝 Poe Quora's model router — one points balance spends across 200+ bots
✨ Genspark Super Agent over 9 LLMs and 80+ tools, plus its own persistent "Claw" agent
🔎 Perplexity Max The $200/mo tier that turns Perplexity into a model-council aggregator
Code & automate
Claude Code Terminal agent that reads, writes, and runs your codebase
🏗 Building complex codebases A repeatable workflow for shipping features in a large codebase with an AI agent
🍵 Gitea The self-hosted GitHub — own the server, own the code
🐙 GitHub Where the world's code lives and ships
🔁 Test-Driven Prompt Engineering Write your examples first, then edit the prompt until they pass
🚀 From demo to production agent What separates a real agent from a demo
🤖 AI for robotics & edge devices The same AI assistants, applied to firmware, robots and constrained hardware
🧩 Skills, tools & extensions The extensibility layer every agent shares — skills, tools, subagents & hooks
Claude Cowork Shared skills and plugins for Claude Code teams
Hermes Self-hosted AI agent that remembers you and writes its own skills
Antigravity Agent-first dev platform — orchestrate many AI agents at once
openclaw Run AI agents on your own devices via WhatsApp or Telegram
opencode Open-source terminal agent — your key, any model, full control
🎙️ opencode + voice Talk to your terminal coding agent — local speech in, local speech out
jcode Open-source terminal harness — any model, even local
Codex OpenAI's terminal coding agent — reads, edits, runs your repo
Claude Agent SDK Build your own agents on the engine behind Claude Code
Claude API Call Claude models from your own code, billed per token
Ollama Run open models locally from the CLI + a local API
Cursor VS Code reimagined with an AI agent that writes and runs your code
GitHub Copilot AI pair programmer built into your editor — writes code as you type
Devin Desktop AI-native IDE whose agent codes whole tasks alongside you
Aider Open-source AI pair programmer in your terminal
⚡ vLLM The serving engine built for many people hitting one model at once
🧱 llama.cpp The engine underneath Ollama — full control, runs on almost anything
🧩 LocalAI One OpenAI-compatible door in front of every engine you run
🔁 llama-swap Hot-swap models on one GPU instead of keeping them all loaded
No-code builders
Base44 No-code app builder with built-in database and hosting included
🐳 Docker Containers, images, and volumes — the mental model before you install anything
🐳 Docker on Windows Install Docker Desktop on Windows via WSL2 and run your first container
🐳 Docker on macOS Install Docker Desktop on Mac (Apple Silicon or Intel) and run your first container
🐳 Docker on Linux Install Docker Engine natively on Linux, run without sudo, and start your first container
⚖ Private AI for your org: buy, build & govern The data-residency spectrum + the governance checklist an institution asks
🇩🇪 IONOS AI Model Hub German-hosted, OpenAI-compatible inference API — ISO 27001 data centers
🇫🇷 Scaleway French cloud — pay-per-token Generative APIs, or rent a GPU by the hour
🇫🇷 OVHcloud Europe's largest independent cloud — serverless AI Endpoints, EU-hosted
Lovable Chat to build full-stack React apps with GitHub export
v0 Prompt full-stack React apps — live preview, one-click deploy on Vercel
Make Visual no-code workflow automation
Zapier Connect 9,000+ apps, no code needed
KNIME Visual data science — no code needed
Flowise Build AI agents visually — no code
Langflow Drag-and-drop builder for AI agents & RAG
CrewAI Python framework for teams of AI agents
🎙️ ElevenLabs The most human-sounding voice AI, plus cloning and dubbing
⚡ Cartesia (Sonic) Sub-100ms streaming TTS built for real-time voice agents
🔊 OpenAI TTS The cheapest good cloud voice, bundled with the GPT ecosystem
📝 Deepgram (Nova-3) The most accurate cloud speech-to-text, batch or live
🎧 Whisper (open source) OpenAI's open-weight speech-to-text, running free on your laptop
🗣️ Kokoro 82M (open source) Tiny open TTS model that sounds far bigger than its 82M params
🎧 OpenAI Realtime API Managed speech-to-speech API — talk to gpt-realtime directly
🔵 Google Gemini Live Cheapest full-stack managed realtime voice, unified per-token billing
🎙️ ElevenLabs Conversational AI Best-in-class voice quality via STT → LLM → ElevenLabs TTS pipeline
☎️ Retell AI Transparent bring-your-own-API voice agent platform, popular for phone agents
🔧 Pipecat Open-source, pluggable voice pipeline: VAD → STT → LLM → TTS
🇫🇷 Moshi (Kyutai) Open-source speech-native model — full-duplex, sub-200ms, EU-hosted
🧑💼 HeyGen Market-leading talking-head avatars — pre-recorded and realtime
🎬 Synthesia Enterprise training-video avatars with PowerPoint-to-video and SCORM export
🗣️ D-ID Photo-to-talking-head plus sub-0.5s realtime LLM-connected agents
👤 Tavus (CVI) The realtime interactive avatar specialist — low-latency conversational video
🦾 SadTalker Open-source single-photo talking head — free, self-hosted, needs a GPU
💋 LatentSync / Wav2Lip Open-source lip-sync for dubbing existing video into another language
🎥 Sora 2 Cinema-grade text-to-video with character consistency
🎞️ Google Veo 3 Fast, budget text-to-video with native audio
🎬 Runway Gen-4 Pro filmmaking tools: motion control, multi-shot consistency
🐉 Kling Physics-accurate video at the lowest proprietary cost
🌀 Wan 2.2 Open-source (Apache-2.0), the most versatile self-hosted video model
🎇 HunyuanVideo Large open text-to-video model, needs a serious GPU
🎼 LTX-Video Fast open video — first OSS model with synced audio in one pass
🖼️ GPT Image 2 OpenAI's flagship image model — sharp text, 2K output, ChatGPT-native
🎨 Midjourney V8.1 The aesthetic leader — subscription only, no free tier, no API
🔓 FLUX.2 The open one — Apache-2.0 weights up to a frontier-grade paid API
🍌 Nano Banana Pro (Gemini 3 Pro Image) Google's multimodal reasoning image model — grounded, accurate text and data
🕸️ Stable Diffusion 3.5 + ComfyUI Fully local, fully free — the open node-graph workflow
🔥 Adobe Firefly (Image Model 5) The commercially-safe choice — trained on licensed data, IP-indemnified
🚀 Dokploy Self-hosted Vercel/Heroku alternative — git push, auto-deploy, your own server
3See also
💬 Discuss this chapter
Ask, share, or report — over on the Heidelberg AI community forum.