Skip to content

Generative AI & LLMs

Jagged technological frontier

AI being good at some tasks while failing at others that seem equally easy or similar.

Example

An assistant handles one analysis well but makes basic mistakes on a closely related problem.

Why people use it

It explains why success on one task does not prove AI will manage a similar one.

What you'll hear

“It handled the hard question but missed the easy one.”

What this means for you

Test each important use case rather than extrapolating from impressive examples.

Can you control it?

No

No direct control. This describes a wider issue, concept or result rather than something you can simply switch on or off in a tool.

Common questions

Can success on a hard task establish reliability on easier-looking tasks?
No. Capability does not increase uniformly across tasks.
Does this only happen with new AI systems?
No. Even familiar systems can have uneven strengths across different tasks.
Can a broad score hide the unevenness?
Yes. An average can conceal particular tasks where the system performs poorly.

Related terms

Still have questions?

Up to 500 characters.

Ask LATHIC about AI. Relevant glossary entries may be included.

Your question, the glossary entries it matches, and a rotating pseudonymous identifier go to Microsoft Azure’s OpenAI service through Vercel AI Gateway to generate an answer. Zero retention and no training are required of the provider, and LATHIC does not save your question or answer. Privacy Notice