Skip to main content
LIVE SUN, 4 OCT, 2026 BENGALURU · 28°C EDITION № 157 · FREE · NO LOGIN
AI AI · 2 MIN READ

AI expert Jay Kruer remains bearish on large language models' autonomy

Jay Kruer, an AI researcher, expressed skepticism about the current capabilities of large language models (LLMs) in a detailed post published on September 15.

Jay Kruer, an AI researcher, expressed skepticism about the current capabilities of large language models (LLMs) in a detailed post published on September 15. He argued that despite high valuations of frontier AI labs, these models still require extensive human oversight and are far from being fully autonomous replacements for knowledge workers, according to dank.systems.

Kruer outlined several points to support his view, noting that while LLMs perform well on specific tasks they were trained on, their generalization ability is limited. He highlighted issues such as reward hacking and failure under small task variations, emphasizing that current models need rigorous specification to overcome these problems. Kruer also referenced examples like Navier-Stokes simulations and software exploits to illustrate the gap between hype and practical autonomy.

This perspective challenges the prevailing narrative in the AI sector, where many frontier labs are valued based on expectations of near-term breakthroughs in automation. Kruer's critique underscores ongoing limitations in LLMs, contrasting with the industry's optimistic projections. His analysis aligns with observations that companies continue to employ human engineers to supervise AI outputs, reflecting the models' current dependency on human intervention.

Kruer's post thanked several contributors for feedback and serves as a cautionary note on the state of AI autonomy. The detailed critique was published on September 15 on dank.systems, providing a measured assessment of LLM capabilities amid widespread enthusiasm.

Editorial standards. Reported and edited at Startupniti's news desk from the sources listed in the right rail. Every fact traces to a citation. If something looks wrong, write to corrections.
▸ WIRE
Premium content free for first 12 months · sign up to unlock Razorpay subscriptions launch Jan 2027 — ₹199/mo or ₹999/yr Every story reads every Indian tech source so you don't have to Every article cited · trust the source, not just the byline India's startup desk, edited daily Founders · Funding · Policy · Tech — three crawls a day Premium content free for first 12 months · sign up to unlock Razorpay subscriptions launch Jan 2027 — ₹199/mo or ₹999/yr Every story reads every Indian tech source so you don't have to Every article cited · trust the source, not just the byline India's startup desk, edited daily Founders · Funding · Policy · Tech — three crawls a day