AI Atlas: Which AI fits which task?
45 task categories, 95 products and model families, 123 editorial ratings. Build a shortlist, then test it against your own task.
How to read the atlas
Research snapshot: 19 September 2026. Based on documented vendor features, not hands-on tests of every product. F = feature fit (50%), C = controllability (25%), W = workflow integration (25%), each rated 1–5. The weighted total is rounded to half stars. Five stars mean a preferred shortlist for that task, not a universal winner. Output quality, speed and cost per usable result were not measured here. An export block may override the overall score: Udio receives 1/5 for external music delivery.
Medical, legal, HR, financial, construction and industrial uses require professional validation and appropriate human responsibility.
Choose a task
General AI Assistants & LLMs
Take ChatGPT and Claude into the shortlist together; Gemini in particular where Google workflows are already in place. No intelligence winner was measured here. Comparison test: Give every candidate an identical brief with sources, contradictions and the desired output file.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| ChatGPT | 5.0 | 5 / 4 / 5 |
| Claude | 5.0 | 5 / 4 / 5 |
| Gemini / Gemini API | 4.5 | 5 / 4 / 4 |
| Mistral Vibe | 4.0 | 4 / 4 / 4 |
| Grok | 4.0 | 4 / 3 / 4 |
| DeepSeek API | 4.0 | 4 / 4 / 3 |
| Qwen | 4.0 | 4 / 4 / 3 |
Web Research & Sourcing
Perplexity for searching the open web; Gemini Notebook for a defined set of your own sources. The recommendations address different tasks. Comparison test: Require ten verifiable statements; open every original source and count the correct supporting passages.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Perplexity | 4.5 | 5 / 4 / 4 |
| Gemini Notebook / NotebookLM | 5.0 | 5 / 5 / 4 |
| ChatGPT | 4.0 | 4 / 4 / 4 |
Science & Literature Review
Elicit for structured extraction and review workflows; Consensus for literature questions and citation trails. Comparison test: Provide a known set of studies; count missing studies, incorrect figures and unsupported conclusions.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Elicit | 5.0 | 5 / 5 / 4 |
| Consensus | 4.5 | 5 / 4 / 4 |
Programming & Repository Work
Test Codex and Claude Code on well-scoped project tasks; Cursor for work in an AI editor; Copilot for a GitHub-centric team. Comparison test: Have three real issues solved with hidden regression tests; only reviewed changes count.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| OpenAI Codex | 5.0 | 5 / 5 / 5 |
| Claude Code | 5.0 | 5 / 5 / 5 |
| Cursor | 5.0 | 5 / 5 / 5 |
| GitHub Copilot | 5.0 | 5 / 4 / 5 |
Websites, Apps & Prototypes
v0 for visually controlled web interfaces; Lovable or Replit for an integrated entry into app building. No automatic seal of production readiness. Comparison test: Test login, role separation, mobile view, export and restore on the same mini project.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| v0 | 5.0 | 5 / 5 / 5 |
| Lovable | 4.5 | 5 / 4 / 4 |
| Replit Agent | 5.0 | 5 / 4 / 5 |
Image Generation & Visual Concepts
Shortlist Midjourney for finding a style; FLUX for image workflows you can integrate; Firefly where Adobe work already exists. The aesthetic winner remains open. Comparison test: Generate the same six subjects several times; judge detail, identity, text and the rework required.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Midjourney | 4.5 | 5 / 4 / 3 |
| FLUX / Black Forest Labs | 5.0 | 5 / 5 / 5 |
| Adobe Firefly | 5.0 | 5 / 5 / 5 |
| ChatGPT | 4.0 | 4 / 4 / 4 |
| Ideogram | 5.0 | 5 / 5 / 4 |
| Google Imagen | 4.5 | 5 / 4 / 4 |
Design, Advertising & Presentations
Canva for recurring brand formats, Gamma for a first structured slide deck. Final acceptance happens on the exported document. Comparison test: Translate one shared branding brief into ten slides and three social formats.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Canva AI | 5.0 | 5 / 5 / 5 |
| Gamma | 4.5 | 5 / 4 / 4 |
| Adobe Firefly | 4.5 | 4 / 5 / 5 |
Cinematic Video & Generated Scenes
For our own production, test Higgsfield as the working interface plus a deliberate choice of model. Runway as an alternative production environment; Veo for scenes with audio; Seedance for reference-heavy specifications. No blind-test winner. Comparison test: Generate three connected shots of the same person: close-up, walking movement, interaction. Record continuity and the cost of every failed attempt.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Higgsfield | 5.0 | 5 / 5 / 4 |
| Runway | 5.0 | 5 / 5 / 5 |
| Google Veo | 4.5 | 5 / 4 / 4 |
| Dreamina Seedance | 5.0 | 5 / 5 / 4 |
| Kling AI | 4.0 | 4 / 4 / 3 |
| Luma | 4.0 | 4 / 4 / 4 |
| Sora | 1.0 | 1 / 1 / 1 |
Avatars, Presenters & Training Videos
HeyGen for multilingual presenters; Synthesia for standardized corporate training. Neither automatically replaces a staged film production. Comparison test: Sign off the same 60-second script with names, figures and gestures in German and Vietnamese.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| HeyGen | 5.0 | 5 / 4 / 5 |
| Synthesia | 5.0 | 5 / 4 / 5 |
Video Editing & Full Film Finishing
Resolve or Premiere for the complete film; CapCut for short social formats; Descript for speech-driven content. Comparison test: Fully export one project with ten clips, music, subtitles and two formats.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| DaVinci Resolve | 5.0 | 5 / 5 / 5 |
| Adobe Premiere | 5.0 | 5 / 5 / 5 |
| CapCut | 4.5 | 5 / 4 / 4 |
| Descript | 4.5 | 5 / 4 / 4 |
Shorts, Subtitles & Repurposing
OpusClip for suggestions out of long recordings; CapCut for targeted rework. A virality score is not a guarantee of reach. Comparison test: Pull three clips from an interview; check context fidelity, the opening seconds, subtitles and safe areas.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| OpusClip | 4.5 | 5 / 4 / 4 |
| CapCut | 5.0 | 5 / 5 / 4 |
| Descript | 4.0 | 4 / 4 / 4 |
Video Restoration & Upscaling
Topaz as the specialist candidate, Resolve as the alternative inside an existing edit project. Not every subject benefits from generated detail. Comparison test: Compare faces, text and fast motion before and after processing at normal playback speed.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Topaz Video | 5.0 | 5 / 5 / 4 |
| DaVinci Resolve | 4.5 | 4 / 4 / 5 |
Songs, Rap & Music with Vocals
Suno is our first choice for a complete rap with given lyrics. Test Google Lyria and Eleven Music as counter-candidates. Udio receives only one star for this purpose because of the documented export obstacle. Comparison test: Use the same eight lines and style specifications; check lyric fidelity, pronunciation, hook, song ending and downloadable files.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Suno | 5.0 | 5 / 5 / 5 |
| Eleven Music | 4.0 | 4 / 4 / 4 |
| Google Lyria | 4.5 | 5 / 4 / 4 |
| Udio | 1.0 | 3 / 3 / 1 |
Background Music, Instrumentals & Sound Design
SOUNDRAW for adjustable backgrounds; AIVA for compositional work and MIDI; Stable Audio for your own audio pipelines. Comparison test: Score a 30-second commercial: the intended arc of tension, a clean ending and room for speech.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| SOUNDRAW | 5.0 | 5 / 5 / 4 |
| AIVA | 5.0 | 5 / 5 / 4 |
| Stable Audio | 4.0 | 4 / 5 / 3 |
| Eleven Music | 4.0 | 4 / 4 / 4 |
Voiceover & Synthetic Voices
Test ElevenLabs first for standalone voiceover; HeyGen when a presenter is needed at the same time. Comparison test: The same text with proper names, figures and emotion; a native speaker judges intelligibility and naturalness.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| ElevenLabs Text to Speech | 5.0 | 5 / 5 / 5 |
| HeyGen | 4.0 | 4 / 4 / 4 |
Video Translation & Dubbing
HeyGen where lip sync is wanted, Rask as the specialized alternative. Source and translation have to carry the same message. Comparison test: Localize a two-speaker video with technical terms; check timing, names and speaker changes.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| HeyGen | 5.0 | 5 / 4 / 5 |
| Rask AI | 4.5 | 5 / 4 / 4 |
Text Translation & International Content
DeepL for focused translation; Claude for editorial adaptation and for explaining alternatives. Distinguish translation from free advertising adaptation. Comparison test: Compare German, English and Vietnamese with figures, forms of politeness and a binding glossary.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| DeepL | 5.0 | 5 / 4 / 5 |
| Claude | 4.5 | 4 / 5 / 4 |
Meetings, Transcription & Tasks
Fathom for compact meeting notes, Fireflies for downstream team and CRM work, Otter as a further candidate. The actual meeting language is part of the decision. Comparison test: Review one approved recording with three speakers: names, figures, decisions and tasks.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Fathom | 4.5 | 5 / 4 / 4 |
| Fireflies | 5.0 | 5 / 4 / 5 |
| Otter | 4.5 | 5 / 4 / 4 |
Office, Email & Internal Knowledge
The existing work environment decides: Microsoft 365, Google Workspace or Notion. Switching for the AI alone can create more effort than benefit. Comparison test: Carry out one task involving mail, a document and an appointment; deliberately vary permissions and the age of the sources.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Microsoft 365 Copilot | 5.0 | 5 / 4 / 5 |
| Gemini in Google Workspace | 5.0 | 5 / 4 / 5 |
| Notion AI | 5.0 | 5 / 5 / 4 |
Data Analysis, Spreadsheets & Reports
ChatGPT for combined analysis and reporting; Julius as the specialized alternative; Copilot for an Excel-centric team. Computational accuracy has not been measured yet. Comparison test: Analyze a spreadsheet with missing values, duplicates and a deliberately misleading correlation.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| ChatGPT | 5.0 | 5 / 4 / 5 |
| Julius AI | 4.0 | 4 / 4 / 4 |
| Microsoft 365 Copilot | 5.0 | 5 / 4 / 5 |
Automation & Cross-System Agents
n8n for technically supervised custom flows, Make for visual integration, Zapier Agents for connected SaaS tasks. The model is only one component. Comparison test: Process an inquiry into a CRM draft; test duplicates, an API outage, retries and human approval.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| n8n | 5.0 | 5 / 5 / 5 |
| Make | 5.0 | 5 / 5 / 5 |
| Zapier Agents | 5.0 | 5 / 4 / 5 |
Customer Service, Chatbots & Conversational Agents
Fin where a support organization already exists; Voiceflow for individually designed dialogues; Copilot Studio in a Microsoft enterprise. Comparison test: 30 inquiries including an unknown answer, a complaint and a data change; count correct escalations.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Fin / Intercom | 5.0 | 5 / 4 / 5 |
| Voiceflow | 5.0 | 5 / 5 / 5 |
| Microsoft Copilot Studio | 5.0 | 5 / 5 / 5 |
Sales, CRM & Customer Processes
Look at HubSpot first if it is already the CRM; Agentforce with Salesforce. No CRM switch for the sake of an agent name alone. Comparison test: Real anonymized sample data: summarize the need, propose the next step, avoid duplicates.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| HubSpot Agent Hub | 5.0 | 5 / 4 / 5 |
| Salesforce Agentforce | 5.0 | 5 / 5 / 5 |
Marketing Copy, Content & SEO/GEO
Jasper for brand-governed team production; Surfer for content planning and diagnostics; Claude as a flexible editorial assistant. No tool replaces your own research. Comparison test: Produce one sourced article; assess facts, originality, search intent and the manual rework needed.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Jasper | 5.0 | 5 / 5 / 4 |
| Surfer | 4.5 | 5 / 4 / 4 |
| Claude | 4.5 | 4 / 5 / 4 |
E-Commerce & Product Images
Photoroom for product-oriented image workflows; Canva for the ad formats built from them. Comparison test: Ten real products: identical colors, shape and labeling, different backgrounds and formats.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Photoroom | 5.0 | 5 / 5 / 5 |
| Canva AI | 4.5 | 4 / 5 / 5 |
3D, Game Assets & Animation
Compare Meshy and Tripo using the same references. Blender remains a separate tool for cleanup and staging, not a comparable AI model candidate. Comparison test: Generate the same object; check geometry, texture, export, rig and behavior in the target program.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Meshy | 5.0 | 5 / 5 / 5 |
| Tripo | 4.5 | 5 / 4 / 4 |
Local AI & Self-Hosted Model Environments
LM Studio for an accessible local entry point; Ollama for a technically integrated model runtime. Neither is the language model itself. Comparison test: Run the same specific model on the same hardware; compare memory, time, network traffic and result.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| LM Studio | 5.0 | 5 / 5 / 4 |
| Ollama | 5.0 | 5 / 5 / 5 |
OCR, Invoices & Document Processes
Choose between Document AI and Textract according to the cloud already in use and the document type. An extracted value is not yet an approved booking. Comparison test: 100 approved sample pages: check the fields against target values; route low-confidence cases to a human.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Google Document AI | 5.0 | 5 / 5 / 5 |
| Amazon Textract | 5.0 | 5 / 4 / 5 |
Learning, Training & Knowledge Transfer
Khanmigo for pedagogically designed support; Gemini Notebook for learning from your own documents. Comparison test: Have the learner explain first, then give a hint; measure the learning outcome with new exercises.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Khanmigo | 4.5 | 5 / 4 / 4 |
| Gemini Notebook / NotebookLM | 5.0 | 5 / 5 / 4 |
Specialist Legal Work
Harvey is a specialist-software candidate for professionally supervised workflows. The stars rate only the documented workflow, not legal correctness. Comparison test: A lawyer reviews identical cases, citations, currency and missing counterarguments.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Harvey | 4.0 | 4 / 4 / 4 |
Security & Media Authenticity
Security Copilot for security work in a matching enterprise stack; Resemble for media authenticity. Do not read the differing purposes as a head-to-head. Comparison test: Use known genuine and manipulated cases; count false alarms and missed cases separately.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Microsoft Security Copilot | 4.5 | 4 / 4 / 5 |
| Resemble AI | 4.0 | 4 / 4 / 4 |
Industry, Simulation & Physical AI
Omniverse as a platform candidate; the choice requires a technical proof of concept of your own, not a general chatbot ranking. Comparison test: Scope one real plant or robotics task with defined simulation and real-world tests.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| NVIDIA Omniverse | 4.5 | 4 / 5 / 4 |
Podcasts & Audio Restoration
Adobe Podcast for speech cleanup; Descript for editorial cutting. Compare the same recording, not demos. Comparison test: Process one of your own recordings with background noise; judge intelligibility, artifacts and naturalness blind.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Adobe Podcast | 4.5 | 5 / 4 / 4 |
| Descript | 5.0 | 5 / 5 / 5 |
Architecture, Real Estate & Construction Planning
Forma for early site and building planning. Image generators are suited to mood, not to reliable dimensions. Comparison test: Assess one plot with real constraints; document deviations, the data basis and the BIM handover.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Autodesk Forma | 4.5 | 5 / 4 / 4 |
| Adobe Firefly | 3.5 | 3 / 4 / 4 |
Medical Documentation
Dragon Copilot is a candidate for professionally supervised documentation workflows. Not a diagnostics ranking. Comparison test: Clinical staff compare permitted test conversations with the notes: omissions, incorrect entries and correction time.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Microsoft Dragon Copilot | 4.0 | 4 / 4 / 4 |
People, HR & Recruiting Processes
Workday AI in a Workday environment; Joule with SAP. Choose the systems according to the existing process landscape. Comparison test: Test one administrative request from intake to the correct answer; do not rate automated candidate decisions.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Workday AI | 5.0 | 5 / 4 / 5 |
| SAP Joule Assistants | 5.0 | 5 / 4 / 5 |
Finance, Procurement & ERP
Compare SAP Joule and Workday according to the modules in place. Extraction, analysis and approval remain separate steps. Comparison test: Test one procurement or finance case against approval rules and target values; count incorrect postings separately.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| SAP Joule Assistants | 5.0 | 5 / 4 / 5 |
| Workday AI | 5.0 | 5 / 4 / 5 |
| Amazon Textract | 3.5 | 3 / 4 / 4 |
Enterprise Search & Internal Knowledge
Glean for a dedicated enterprise search; Notion and Microsoft where the existing knowledge landscape fits. Comparison test: Ten questions with permitted and restricted sources: check the correct source passage, currency and permission boundaries.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Glean | 5.0 | 5 / 4 / 5 |
| Notion AI | 5.0 | 5 / 4 / 5 |
| Microsoft 365 Copilot | 5.0 | 5 / 4 / 5 |
Shop Search, Retrieval & Recommendations
Algolia as the search platform; Cohere as a reranking component. Built for different architectures, not interchangeable one for one. Comparison test: Evaluate 50 real search queries: correct hits in the top five, zero-result rates and response time.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Algolia AI Search | 5.0 | 5 / 5 / 5 |
| Cohere Rerank | 4.5 | 5 / 4 / 4 |
Forecasting & Custom ML Applications
Choose between Databricks and DataRobot according to data platform and operations. Do not infer forecast quality from product descriptions. Comparison test: Use time-separated training and test data; compare against a simple historical baseline forecast.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Databricks AI | 5.0 | 5 / 5 / 5 |
| DataRobot | 4.5 | 5 / 4 / 4 |
Science, Climate & Biological Models
AlphaFold for molecular structures; WeatherNext for weather. These are fields with completely different target variables. Comparison test: Biology against independently verified structures; weather against held-out periods and regional reference data.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| AlphaFold | 4.5 | 5 / 4 / 4 |
| WeatherNext | 4.0 | 5 / 3 / 3 |
Robotics, Logistics & Visual Inspection
Gemini Robotics for model-based robotics, Omniverse for simulation and Roboflow for visual recognition. Building blocks instead of a universal robot. Comparison test: Test a defined real-world task including unknown cases; measure incorrect actions, failures and recovery.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Gemini Robotics | 4.0 | 4 / 4 / 3 |
| NVIDIA Omniverse | 5.0 | 5 / 5 / 4 |
| Roboflow | 5.0 | 5 / 5 / 4 |
Telephony & Interactive Voice Agents
ElevenLabs Agents for voice dialogues; Voiceflow for designed dialogues with integrations. Sign off the complete customer case first. Comparison test: A call with an interruption, an accent, a wrong customer number and a human handover; allow no invented commitments.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| ElevenLabs Agents | 4.5 | 5 / 4 / 4 |
| Voiceflow | 5.0 | 5 / 5 / 4 |
Testing, Monitoring & Operating Agents
LangSmith for observability; Databricks for AI inside the data platform. Both need target cases and error thresholds of their own. Comparison test: Inject known faults: wrong source, timeout, duplicate job, unauthorized tool; check detection and recovery.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| LangSmith | 5.0 | 5 / 5 / 5 |
| Databricks AI | 4.5 | 4 / 5 / 5 |
Travel, Visas & Multilingual Service
Connect document capture, translation and workflow. For Vietnam visas, correct data and reliable handovers are what count. Comparison test: Run approved sample cases with deliberately contradictory data; update official requirements separately.
| Product / vendor source | Editorial fit / 5 | F / C / W |
|---|---|---|
| Google Document AI | 5.0 | 5 / 5 / 4 |
| DeepL | 4.5 | 5 / 4 / 4 |
| n8n | 5.0 | 5 / 5 / 5 |
Complete source material
This compact overview includes every category and rating. The German original documents add use cases, limitations and cost factors for each product. Check current prices, licences, access and regional availability with the vendor before deciding.
Our practical test: customer enquiries (PDF, German)
A separate executed model test: ten synthetic enquiries, three models and 30 runs. It is not a customer case or evidence for every product in the atlas.
Our practical test: customer enquiries (PDF, German)Start potential analysis
If you want to prioritize a real process, a few clear inputs are enough for a strong first assessment.