BS Summary: This article contains 26 faulty reasoning types, including Appeal to Authority, Hasty Generalization, and Confirmation Bias, with Negativity Bias as the most egregious example at 35.6% saturation with 222 hits. Analysis detected 1,417 faulty-reasoning hits from 624 analyzed words, generating a BS Score of 57.3% and a BS Rank of 62% (8,154 of 21,147 articles). This article is worse (more manipulative) than 61.40% of the article peer group.

To date, AI industry spending has topped $1.6 trillion, and shows no sign of slowing anytime soon. 
So what do we actually have to show for it? 
Historically, it’s been a whole lot of nothing: as numerous studies have shown us, tools like AI chatbots and autonomous agents have been ineffective at completing real world tasks in a competent way. 
The tech industry insists that’s all about to change within the next few years , as AI’s capabilities grow by leaps and bounds, enabling economic growth the likes of which the world has never seen. 
But is it really ? 
Not necessarily. 
A new study out of the University of California Berkeley’s Center for Responsible, Decentralized Intelligence  flagged by the College Fix  shows that frontier AI tools of all makes and models are still incapable of completing the vast majority of workplace tasks at an acceptable level, throwing a major wrench in the tech industry’s assertions that the AI revolution is imminent. 
To come to that conclusion, the UC researchers designed a rigorous assessment they call the “Agents’ Last Exam,” developed to test “job-readiness” across numerous state-of-the-art AI models. 
Basically, the ALE  an impish riff on  Humanity’s Last Exam   is designed to put an AI system through its paces, covering “more than 1,500 expert-sourced tasks spanning 55 occupations,” the researchers wrote in a press release . 
Those test spans the typical line-up of AI-exposed jobs like software engineering and graphic design, but also a substantial number of jobs whose fates remain less certain, such as maritime engineering, agriculture, audio production, and public health operations. 
Using the ALE benchmark , researchers took a hard look at advanced “closed” models  proprietary AI systems developed by private companies  like Anthropic’s Fable 5, OpenAI’s GPT-5.5, Cursor’s Composer 2.5, and Google’s Gemini 3.1 Pro. 
(For good measure, they also looked at two open-source models by Chinese developers.) 
As cutting-edge as these AI models are, the research found that they’re far from ready for the complex needs of the modern workplace. 
Out of all of the models put through the gauntlet, each of them failed spectacularly. 
OpenAI’s GPT-5.5 came in with the highest score: a passing rate of just 24 percent overall. 
“Today’s agents can solve a meaningful fraction of professional tasks,” the researchers wrote. 
“However, when we look at the hardest tasks that require sustained reasoning, deep domain expertise, and reliable execution over long horizons, they are still far from human-level performance.” 
And as tasks became more complicated, even those meager aggregate scores fell off fast. 
“On ALE’s hardest tier, every frontier agent we tested, including Fable 5, achieved a 0 percent success rate,” the presser explains. 
The researchers also break down some cost considerations. 
The cutting edge Fable 5, they note, delivers “similar performance” to models like GPT-5.5 and Composer 2.5, “while costing roughly 4-12× more per completed task.” 
Despite the horrible test results, researchers caution that the technology could still upend the job market for more AI-exposed  as plenty of corporate executives have shown us , the tech doesn’t need to work particularly well to keep workers on their back heels. 
“Even if current pass rates remain relatively low, occupations dominated by routine and well-defined procedures are likely to experience disruption first, while decision-intensive roles will remain more resilient for longer,” Berkeley computer science researcher and study co-author Dawn Song told College Fix . 
“The key factor,” Song added, “is not the industry itself, but the nature of the work.” 
More on AI: OpenAI Appears to Be Missing Its Sales Goals by a Vast Margin 
The post Frontier AI Is Faceplanting at Real-World Workplace Tasks appeared first on Futurism . 
Article reasoning-pattern comparisonThis article: 17.9%Joe Wilkins: 6.6%Futurism: 6.0%Confirmation Bias17.9%This article: 2.6%Joe Wilkins: 1.6%Futurism: 1.6%Anchoring Bias2.6%This article: 6.7%Joe Wilkins: 6.3%Futurism: 6.2%Availability Heuristic6.7%This article: 6.1%Joe Wilkins: 1.6%Futurism: 1.5%Representativeness Heuristic6.1%This article: 0.0%Joe Wilkins: 0.3%Futurism: 0.6%Hindsight Bias0.0%This article: 0.0%Joe Wilkins: 3.7%Futurism: 2.6%Overconfidence Bias0.0%This article: 1.3%Joe Wilkins: 13.2%Futurism: 11.5%Framing Effect1.3%This article: 6.9%Joe Wilkins: 1.1%Futurism: 1.2%Loss Aversion6.9%This article: 2.6%Joe Wilkins: 0.7%Futurism: 0.6%Status Quo Bias2.6%This article: 0.0%Joe Wilkins: 0.1%Futurism: 0.2%Sunk Cost Effect0.0%This article: 5.6%Joe Wilkins: 2.0%Futurism: 2.2%Optimism Bias5.6%This article: 6.1%Joe Wilkins: 4.5%Futurism: 4.3%Pessimism Bias6.1%This article: 35.6%Joe Wilkins: 25.5%Futurism: 22.6%Negativity Bias35.6%This article: 0.0%Joe Wilkins: 1.6%Futurism: 1.6%Self-Serving Bias0.0%This article: 7.1%Joe Wilkins: 1.4%Futurism: 1.1%Fundamental Attribution Error7.1%This article: 0.0%Joe Wilkins: 0.2%Futurism: 0.2%Actor-Observer Bias0.0%This article: 0.0%Joe Wilkins: 1.2%Futurism: 0.6%In-Group Bias0.0%This article: 0.0%Joe Wilkins: 0.4%Futurism: 0.3%Out-Group Homogeneity Bias0.0%This article: 0.0%Joe Wilkins: 1.3%Futurism: 1.1%Halo Effect0.0%This article: 0.0%Joe Wilkins: 0.7%Futurism: 0.6%Horn Effect0.0%This article: 0.0%Joe Wilkins: 0.1%Futurism: 0.1%Dunning-Kruger Effect0.0%This article: 2.4%Joe Wilkins: 2.2%Futurism: 2.2%Recency Bias2.4%This article: 0.0%Joe Wilkins: 0.2%Futurism: 0.5%Primacy Effect0.0%This article: 0.0%Joe Wilkins: 0.1%Futurism: 0.1%Blind-Spot Bias0.0%This article: 0.0%Joe Wilkins: 1.4%Futurism: 1.3%Ad Hominem0.0%This article: 5.6%Joe Wilkins: 1.1%Futurism: 0.6%Straw Man5.6%This article: 21.8%Joe Wilkins: 5.6%Futurism: 6.8%Appeal to Authority21.8%This article: 2.6%Joe Wilkins: 2.5%Futurism: 2.5%False Dilemma2.6%This article: 0.0%Joe Wilkins: 2.9%Futurism: 2.6%Slippery Slope0.0%This article: 0.0%Joe Wilkins: 0.3%Futurism: 0.2%Circular Reasoning0.0%This article: 18.3%Joe Wilkins: 12.8%Futurism: 10.5%Hasty Generalization18.3%This article: 9.9%Joe Wilkins: 1.0%Futurism: 0.9%Red Herring9.9%This article: 0.0%Joe Wilkins: 1.5%Futurism: 1.2%Bandwagon0.0%This article: 0.0%Joe Wilkins: 11.4%Futurism: 9.3%Appeal to Emotion0.0%This article: 1.6%Joe Wilkins: 1.6%Futurism: 1.5%Begging the Question1.6%This article: 0.0%Joe Wilkins: 3.0%Futurism: 3.9%Post Hoc (False Cause)0.0%This article: 7.1%Joe Wilkins: 0.1%Futurism: 0.2%Tu Quoque7.1%This article: 0.0%Joe Wilkins: 1.2%Futurism: 1.1%Burden of Proof0.0%This article: 0.0%Joe Wilkins: 0.2%Futurism: 0.3%Appeal to Nature0.0%This article: 0.0%Joe Wilkins: 0.2%Futurism: 0.3%Composition/Division0.0%This article: 1.3%Joe Wilkins: 4.3%Futurism: 4.3%Anecdotal1.3%This article: 0.0%Joe Wilkins: 0.0%Futurism: 0.1%No True Scotsman0.0%This article: 13.0%Joe Wilkins: 2.3%Futurism: 2.7%Ambiguity (Equivocation)13.0%This article: 0.0%Joe Wilkins: 0.0%Futurism: 0.0%Gambler’s Fallacy0.0%This article: 0.0%Joe Wilkins: 0.2%Futurism: 0.2%Middle Ground0.0%This article: 0.0%Joe Wilkins: 0.2%Futurism: 0.1%Personal Incredulity0.0%This article: 0.0%Joe Wilkins: 0.4%Futurism: 0.2%Special Pleading0.0%This article: 0.0%Joe Wilkins: 0.6%Futurism: 0.3%Genetic Fallacy0.0%This article: 16.5%Joe Wilkins: 3.4%Futurism: 3.9%Unattributed Quote16.5%This article: 0.8%Joe Wilkins: 2.8%Futurism: 2.6%Quote-first Misdirection0.8%This article: 10.9%Joe Wilkins: 27.1%Futurism: 20.8%Biased Writer Voice10.9%This article: 0.0%Joe Wilkins: 1.7%Futurism: 2.1%Indoctrination0.0%This article: 0.0%Joe Wilkins: 5.2%Futurism: 2.4%Politically Left Leaning Bias0.0%This article: 9.9%Joe Wilkins: 0.4%Futurism: 0.3%Politically Right Leaning Bias9.9%This article: 7.1%Joe Wilkins: 1.5%Futurism: 1.6%Attempt to Sell a Product or S…7.1%

624 words analyzed.

Speakers

1speaker9.5%attributed speech565writer words
Voice mapSelect a segment to jump to its words
Writer's voice • 8 words • 100.0% coverageWriter's voice • 17 words • 0.0% coverageWriter's voice • 10 words • 100.0% coverageWriter's voice • 33 words • 100.0% coverageWriter's voice • 35 words • 0.0% coverageWriter's voice • 5 words • 100.0% coverageWriter's voice • 2 words • 100.0% coverageWriter's voice • 62 words • 100.0% coverageWriter's voice • 27 words • 0.0% coverageWriter's voice • 41 words • 0.0% coverageWriter's voice • 38 words • 0.0% coverageWriter's voice • 37 words • 0.0% coverageWriter's voice • 13 words • 0.0% coverageWriter's voice • 23 words • 0.0% coverageWriter's voice • 15 words • 100.0% coverageWriter's voice • 16 words • 0.0% coverageWriter's voice • 13 words • 100.0% coverageWriter's voice • 28 words • 100.0% coverageWriter's voice • 14 words • 0.0% coverageWriter's voice • 21 words • 0.0% coverageWriter's voice • 8 words • 0.0% coverageWriter's voice • 25 words • 0.0% coverageWriter's voice • 44 words • 100.0% coverageDawn Song • 43 words • 0.0% coverageDawn Song • 16 words • 0.0% coverageWriter's voice • 15 words • 0.0% coverageWriter's voice • 15 words • 0.0% coverage
Selected voice

Dawn Song

100%flagged-word coverage
59 attributed words100% of attributed speech82% writer coverage
0%10.0%20.0%Unattributed Quote-18.2 ptsWriter: 18.2%Dawn Song: 0.0%0.0%Biased Writer Voice-12.0 ptsWriter: 12.0%Dawn Song: 0.0%0.0%Politically Right Leaning -11.0 ptsWriter: 11.0%Dawn Song: 0.0%0.0%Attempt to Sell a Product -7.8 ptsWriter: 7.8%Dawn Song: 0.0%0.0%Quote-first Misdirection-0.9 ptsWriter: 0.9%Dawn Song: 0.0%0.0%

Attribution is sentence-level. Pattern percentages are calculated only from words assigned to that voice.

Loading…
Loading…
Loading…

Analysis

Hover over highlighted words in the article to view the associated bias or fallacy analysis.