AI Atlas: Which AI fits which task?

45 task categories, 95 products and model families, 123 editorial ratings. Build a shortlist, then test it against your own task.

2026-09-24
Illustration: a compass connects different tasks to suitable tool paths.

How to read the atlas

Research snapshot: 19 September 2026. Based on documented vendor features, not hands-on tests of every product. F = feature fit (50%), C = controllability (25%), W = workflow integration (25%), each rated 1–5. The weighted total is rounded to half stars. Five stars mean a preferred shortlist for that task, not a universal winner. Output quality, speed and cost per usable result were not measured here. An export block may override the overall score: Udio receives 1/5 for external music delivery.

Medical, legal, HR, financial, construction and industrial uses require professional validation and appropriate human responsibility.

Choose a task

01

General AI Assistants & LLMs

Take ChatGPT and Claude into the shortlist together; Gemini in particular where Google workflows are already in place. No intelligence winner was measured here. Comparison test: Give every candidate an identical brief with sources, contradictions and the desired output file.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
ChatGPT5.05 / 4 / 5
Claude5.05 / 4 / 5
Gemini / Gemini API4.55 / 4 / 4
Mistral Vibe4.04 / 4 / 4
Grok4.04 / 3 / 4
DeepSeek API4.04 / 4 / 3
Qwen4.04 / 4 / 3
Back to category overview
02

Web Research & Sourcing

Perplexity for searching the open web; Gemini Notebook for a defined set of your own sources. The recommendations address different tasks. Comparison test: Require ten verifiable statements; open every original source and count the correct supporting passages.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Perplexity4.55 / 4 / 4
Gemini Notebook / NotebookLM5.05 / 5 / 4
ChatGPT4.04 / 4 / 4
Back to category overview
03

Science & Literature Review

Elicit for structured extraction and review workflows; Consensus for literature questions and citation trails. Comparison test: Provide a known set of studies; count missing studies, incorrect figures and unsupported conclusions.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Elicit5.05 / 5 / 4
Consensus4.55 / 4 / 4
Back to category overview
04

Programming & Repository Work

Test Codex and Claude Code on well-scoped project tasks; Cursor for work in an AI editor; Copilot for a GitHub-centric team. Comparison test: Have three real issues solved with hidden regression tests; only reviewed changes count.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
OpenAI Codex5.05 / 5 / 5
Claude Code5.05 / 5 / 5
Cursor5.05 / 5 / 5
GitHub Copilot5.05 / 4 / 5
Back to category overview
05

Websites, Apps & Prototypes

v0 for visually controlled web interfaces; Lovable or Replit for an integrated entry into app building. No automatic seal of production readiness. Comparison test: Test login, role separation, mobile view, export and restore on the same mini project.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
v05.05 / 5 / 5
Lovable4.55 / 4 / 4
Replit Agent5.05 / 4 / 5
Back to category overview
06

Image Generation & Visual Concepts

Shortlist Midjourney for finding a style; FLUX for image workflows you can integrate; Firefly where Adobe work already exists. The aesthetic winner remains open. Comparison test: Generate the same six subjects several times; judge detail, identity, text and the rework required.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Midjourney4.55 / 4 / 3
FLUX / Black Forest Labs5.05 / 5 / 5
Adobe Firefly5.05 / 5 / 5
ChatGPT4.04 / 4 / 4
Ideogram5.05 / 5 / 4
Google Imagen4.55 / 4 / 4
Back to category overview
07

Design, Advertising & Presentations

Canva for recurring brand formats, Gamma for a first structured slide deck. Final acceptance happens on the exported document. Comparison test: Translate one shared branding brief into ten slides and three social formats.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Canva AI5.05 / 5 / 5
Gamma4.55 / 4 / 4
Adobe Firefly4.54 / 5 / 5
Back to category overview
08

Cinematic Video & Generated Scenes

For our own production, test Higgsfield as the working interface plus a deliberate choice of model. Runway as an alternative production environment; Veo for scenes with audio; Seedance for reference-heavy specifications. No blind-test winner. Comparison test: Generate three connected shots of the same person: close-up, walking movement, interaction. Record continuity and the cost of every failed attempt.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Higgsfield5.05 / 5 / 4
Runway5.05 / 5 / 5
Google Veo4.55 / 4 / 4
Dreamina Seedance5.05 / 5 / 4
Kling AI4.04 / 4 / 3
Luma4.04 / 4 / 4
Sora1.01 / 1 / 1
Back to category overview
09

Avatars, Presenters & Training Videos

HeyGen for multilingual presenters; Synthesia for standardized corporate training. Neither automatically replaces a staged film production. Comparison test: Sign off the same 60-second script with names, figures and gestures in German and Vietnamese.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
HeyGen5.05 / 4 / 5
Synthesia5.05 / 4 / 5
Back to category overview
10

Video Editing & Full Film Finishing

Resolve or Premiere for the complete film; CapCut for short social formats; Descript for speech-driven content. Comparison test: Fully export one project with ten clips, music, subtitles and two formats.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
DaVinci Resolve5.05 / 5 / 5
Adobe Premiere5.05 / 5 / 5
CapCut4.55 / 4 / 4
Descript4.55 / 4 / 4
Back to category overview
11

Shorts, Subtitles & Repurposing

OpusClip for suggestions out of long recordings; CapCut for targeted rework. A virality score is not a guarantee of reach. Comparison test: Pull three clips from an interview; check context fidelity, the opening seconds, subtitles and safe areas.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
OpusClip4.55 / 4 / 4
CapCut5.05 / 5 / 4
Descript4.04 / 4 / 4
Back to category overview
12

Video Restoration & Upscaling

Topaz as the specialist candidate, Resolve as the alternative inside an existing edit project. Not every subject benefits from generated detail. Comparison test: Compare faces, text and fast motion before and after processing at normal playback speed.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Topaz Video5.05 / 5 / 4
DaVinci Resolve4.54 / 4 / 5
Back to category overview
13

Songs, Rap & Music with Vocals

Suno is our first choice for a complete rap with given lyrics. Test Google Lyria and Eleven Music as counter-candidates. Udio receives only one star for this purpose because of the documented export obstacle. Comparison test: Use the same eight lines and style specifications; check lyric fidelity, pronunciation, hook, song ending and downloadable files.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Suno5.05 / 5 / 5
Eleven Music4.04 / 4 / 4
Google Lyria4.55 / 4 / 4
Udio1.03 / 3 / 1
Back to category overview
14

Background Music, Instrumentals & Sound Design

SOUNDRAW for adjustable backgrounds; AIVA for compositional work and MIDI; Stable Audio for your own audio pipelines. Comparison test: Score a 30-second commercial: the intended arc of tension, a clean ending and room for speech.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
SOUNDRAW5.05 / 5 / 4
AIVA5.05 / 5 / 4
Stable Audio4.04 / 5 / 3
Eleven Music4.04 / 4 / 4
Back to category overview
15

Voiceover & Synthetic Voices

Test ElevenLabs first for standalone voiceover; HeyGen when a presenter is needed at the same time. Comparison test: The same text with proper names, figures and emotion; a native speaker judges intelligibility and naturalness.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
ElevenLabs Text to Speech5.05 / 5 / 5
HeyGen4.04 / 4 / 4
Back to category overview
16

Video Translation & Dubbing

HeyGen where lip sync is wanted, Rask as the specialized alternative. Source and translation have to carry the same message. Comparison test: Localize a two-speaker video with technical terms; check timing, names and speaker changes.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
HeyGen5.05 / 4 / 5
Rask AI4.55 / 4 / 4
Back to category overview
17

Text Translation & International Content

DeepL for focused translation; Claude for editorial adaptation and for explaining alternatives. Distinguish translation from free advertising adaptation. Comparison test: Compare German, English and Vietnamese with figures, forms of politeness and a binding glossary.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
DeepL5.05 / 4 / 5
Claude4.54 / 5 / 4
Back to category overview
18

Meetings, Transcription & Tasks

Fathom for compact meeting notes, Fireflies for downstream team and CRM work, Otter as a further candidate. The actual meeting language is part of the decision. Comparison test: Review one approved recording with three speakers: names, figures, decisions and tasks.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Fathom4.55 / 4 / 4
Fireflies5.05 / 4 / 5
Otter4.55 / 4 / 4
Back to category overview
19

Office, Email & Internal Knowledge

The existing work environment decides: Microsoft 365, Google Workspace or Notion. Switching for the AI alone can create more effort than benefit. Comparison test: Carry out one task involving mail, a document and an appointment; deliberately vary permissions and the age of the sources.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Microsoft 365 Copilot5.05 / 4 / 5
Gemini in Google Workspace5.05 / 4 / 5
Notion AI5.05 / 5 / 4
Back to category overview
20

Data Analysis, Spreadsheets & Reports

ChatGPT for combined analysis and reporting; Julius as the specialized alternative; Copilot for an Excel-centric team. Computational accuracy has not been measured yet. Comparison test: Analyze a spreadsheet with missing values, duplicates and a deliberately misleading correlation.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
ChatGPT5.05 / 4 / 5
Julius AI4.04 / 4 / 4
Microsoft 365 Copilot5.05 / 4 / 5
Back to category overview
21

Automation & Cross-System Agents

n8n for technically supervised custom flows, Make for visual integration, Zapier Agents for connected SaaS tasks. The model is only one component. Comparison test: Process an inquiry into a CRM draft; test duplicates, an API outage, retries and human approval.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
n8n5.05 / 5 / 5
Make5.05 / 5 / 5
Zapier Agents5.05 / 4 / 5
Back to category overview
22

Customer Service, Chatbots & Conversational Agents

Fin where a support organization already exists; Voiceflow for individually designed dialogues; Copilot Studio in a Microsoft enterprise. Comparison test: 30 inquiries including an unknown answer, a complaint and a data change; count correct escalations.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Fin / Intercom5.05 / 4 / 5
Voiceflow5.05 / 5 / 5
Microsoft Copilot Studio5.05 / 5 / 5
Back to category overview
23

Sales, CRM & Customer Processes

Look at HubSpot first if it is already the CRM; Agentforce with Salesforce. No CRM switch for the sake of an agent name alone. Comparison test: Real anonymized sample data: summarize the need, propose the next step, avoid duplicates.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
HubSpot Agent Hub5.05 / 4 / 5
Salesforce Agentforce5.05 / 5 / 5
Back to category overview
24

Marketing Copy, Content & SEO/GEO

Jasper for brand-governed team production; Surfer for content planning and diagnostics; Claude as a flexible editorial assistant. No tool replaces your own research. Comparison test: Produce one sourced article; assess facts, originality, search intent and the manual rework needed.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Jasper5.05 / 5 / 4
Surfer4.55 / 4 / 4
Claude4.54 / 5 / 4
Back to category overview
25

E-Commerce & Product Images

Photoroom for product-oriented image workflows; Canva for the ad formats built from them. Comparison test: Ten real products: identical colors, shape and labeling, different backgrounds and formats.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Photoroom5.05 / 5 / 5
Canva AI4.54 / 5 / 5
Back to category overview
26

3D, Game Assets & Animation

Compare Meshy and Tripo using the same references. Blender remains a separate tool for cleanup and staging, not a comparable AI model candidate. Comparison test: Generate the same object; check geometry, texture, export, rig and behavior in the target program.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Meshy5.05 / 5 / 5
Tripo4.55 / 4 / 4
Back to category overview
27

Local AI & Self-Hosted Model Environments

LM Studio for an accessible local entry point; Ollama for a technically integrated model runtime. Neither is the language model itself. Comparison test: Run the same specific model on the same hardware; compare memory, time, network traffic and result.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
LM Studio5.05 / 5 / 4
Ollama5.05 / 5 / 5
Back to category overview
28

OCR, Invoices & Document Processes

Choose between Document AI and Textract according to the cloud already in use and the document type. An extracted value is not yet an approved booking. Comparison test: 100 approved sample pages: check the fields against target values; route low-confidence cases to a human.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Google Document AI5.05 / 5 / 5
Amazon Textract5.05 / 4 / 5
Back to category overview
29

Learning, Training & Knowledge Transfer

Khanmigo for pedagogically designed support; Gemini Notebook for learning from your own documents. Comparison test: Have the learner explain first, then give a hint; measure the learning outcome with new exercises.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Khanmigo4.55 / 4 / 4
Gemini Notebook / NotebookLM5.05 / 5 / 4
Back to category overview
30

Specialist Legal Work

Harvey is a specialist-software candidate for professionally supervised workflows. The stars rate only the documented workflow, not legal correctness. Comparison test: A lawyer reviews identical cases, citations, currency and missing counterarguments.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Harvey4.04 / 4 / 4
Back to category overview
31

Security & Media Authenticity

Security Copilot for security work in a matching enterprise stack; Resemble for media authenticity. Do not read the differing purposes as a head-to-head. Comparison test: Use known genuine and manipulated cases; count false alarms and missed cases separately.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Microsoft Security Copilot4.54 / 4 / 5
Resemble AI4.04 / 4 / 4
Back to category overview
32

Industry, Simulation & Physical AI

Omniverse as a platform candidate; the choice requires a technical proof of concept of your own, not a general chatbot ranking. Comparison test: Scope one real plant or robotics task with defined simulation and real-world tests.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
NVIDIA Omniverse4.54 / 5 / 4
Back to category overview
33

Podcasts & Audio Restoration

Adobe Podcast for speech cleanup; Descript for editorial cutting. Compare the same recording, not demos. Comparison test: Process one of your own recordings with background noise; judge intelligibility, artifacts and naturalness blind.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Adobe Podcast4.55 / 4 / 4
Descript5.05 / 5 / 5
Back to category overview
34

Architecture, Real Estate & Construction Planning

Forma for early site and building planning. Image generators are suited to mood, not to reliable dimensions. Comparison test: Assess one plot with real constraints; document deviations, the data basis and the BIM handover.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Autodesk Forma4.55 / 4 / 4
Adobe Firefly3.53 / 4 / 4
Back to category overview
35

Medical Documentation

Dragon Copilot is a candidate for professionally supervised documentation workflows. Not a diagnostics ranking. Comparison test: Clinical staff compare permitted test conversations with the notes: omissions, incorrect entries and correction time.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Microsoft Dragon Copilot4.04 / 4 / 4
Back to category overview
36

People, HR & Recruiting Processes

Workday AI in a Workday environment; Joule with SAP. Choose the systems according to the existing process landscape. Comparison test: Test one administrative request from intake to the correct answer; do not rate automated candidate decisions.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Workday AI5.05 / 4 / 5
SAP Joule Assistants5.05 / 4 / 5
Back to category overview
37

Finance, Procurement & ERP

Compare SAP Joule and Workday according to the modules in place. Extraction, analysis and approval remain separate steps. Comparison test: Test one procurement or finance case against approval rules and target values; count incorrect postings separately.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
SAP Joule Assistants5.05 / 4 / 5
Workday AI5.05 / 4 / 5
Amazon Textract3.53 / 4 / 4
Back to category overview
38

Enterprise Search & Internal Knowledge

Glean for a dedicated enterprise search; Notion and Microsoft where the existing knowledge landscape fits. Comparison test: Ten questions with permitted and restricted sources: check the correct source passage, currency and permission boundaries.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Glean5.05 / 4 / 5
Notion AI5.05 / 4 / 5
Microsoft 365 Copilot5.05 / 4 / 5
Back to category overview
39

Shop Search, Retrieval & Recommendations

Algolia as the search platform; Cohere as a reranking component. Built for different architectures, not interchangeable one for one. Comparison test: Evaluate 50 real search queries: correct hits in the top five, zero-result rates and response time.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Algolia AI Search5.05 / 5 / 5
Cohere Rerank4.55 / 4 / 4
Back to category overview
40

Forecasting & Custom ML Applications

Choose between Databricks and DataRobot according to data platform and operations. Do not infer forecast quality from product descriptions. Comparison test: Use time-separated training and test data; compare against a simple historical baseline forecast.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Databricks AI5.05 / 5 / 5
DataRobot4.55 / 4 / 4
Back to category overview
41

Science, Climate & Biological Models

AlphaFold for molecular structures; WeatherNext for weather. These are fields with completely different target variables. Comparison test: Biology against independently verified structures; weather against held-out periods and regional reference data.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
AlphaFold4.55 / 4 / 4
WeatherNext4.05 / 3 / 3
Back to category overview
42

Robotics, Logistics & Visual Inspection

Gemini Robotics for model-based robotics, Omniverse for simulation and Roboflow for visual recognition. Building blocks instead of a universal robot. Comparison test: Test a defined real-world task including unknown cases; measure incorrect actions, failures and recovery.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Gemini Robotics4.04 / 4 / 3
NVIDIA Omniverse5.05 / 5 / 4
Roboflow5.05 / 5 / 4
Back to category overview
43

Telephony & Interactive Voice Agents

ElevenLabs Agents for voice dialogues; Voiceflow for designed dialogues with integrations. Sign off the complete customer case first. Comparison test: A call with an interruption, an accent, a wrong customer number and a human handover; allow no invented commitments.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
ElevenLabs Agents4.55 / 4 / 4
Voiceflow5.05 / 5 / 4
Back to category overview
44

Testing, Monitoring & Operating Agents

LangSmith for observability; Databricks for AI inside the data platform. Both need target cases and error thresholds of their own. Comparison test: Inject known faults: wrong source, timeout, duplicate job, unauthorized tool; check detection and recovery.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
LangSmith5.05 / 5 / 5
Databricks AI4.54 / 5 / 5
Back to category overview
45

Travel, Visas & Multilingual Service

Connect document capture, translation and workflow. For Vietnam visas, correct data and reliable handovers are what count. Comparison test: Run approved sample cases with deliberately contradictory data; update official requirements separately.

Editorial suitability · not a measured ranking
Product / vendor sourceEditorial fit / 5F / C / W
Google Document AI5.05 / 5 / 4
DeepL4.55 / 4 / 4
n8n5.05 / 5 / 5
Back to category overview

Complete source material

This compact overview includes every category and rating. The German original documents add use cases, limitations and cost factors for each product. Check current prices, licences, access and regional availability with the vendor before deciding.

Our practical test: customer enquiries (PDF, German)

A separate executed model test: ten synthetic enquiries, three models and 30 runs. It is not a customer case or evidence for every product in the atlas.

Our practical test: customer enquiries (PDF, German)

Start potential analysis

If you want to prioritize a real process, a few clear inputs are enough for a strong first assessment.

WhatsApp AI assistant