Menu

#Reasoning

576 posts

Feed·
20 of 576 posts
Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things
🖼️
802

Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things

Hacker News·about 2 months ago
#DSU7ejnw

Friday’s big release was Qwen 3.8 27B, an Apache 2 licensed 27B parameter vision-capable LLM from Alibaba’s Qwen research lab. I’ve been looking forward to this one: 27B is an …

15s
Read More
Software Engineering fundamentals matter more than ever
📰
0

Software Engineering fundamentals matter more than ever

Hacker News·about 2 months ago
#8j2fKm3U

The manifestation of my imposter syndrome, for me and today, is what does it mean to be a software engineer. There’s a lot more noise than signal on the Internet about agentic engineering, wh…

15s
Read More
Introducing Muse Glimmer: An Open Agentic Model That Runs on Your Device
🖼️
1202

Introducing Muse Glimmer: An Open Agentic Model That Runs on Your Device

Hacker News·about 2 months ago
#tgOCUNKg
#research#muse#glimmer#model#reasoning#local

Muse Glimmer is a 30-billion-parameter open agentic model from Meta Superintelligence Labs, optimized for always-on local workflows on consumer hardware.

15s
Read More
Is AI Reasoning Right for the Wrong Reasons? | Quanta Magazine
🖼️
0

Is AI Reasoning Right for the Wrong Reasons? | Quanta Magazine

Hacker News·2 months ago
#l2XVyYej

The idea that artificial intelligence can “reason” is more intuitive than ever. But intuitions can be wrong, and the science is far from settled.

15s
Read More
The Session You Cannot Take With You | EARENDIL
📰
0

The Session You Cannot Take With You | EARENDIL

Hacker News·2 months ago
#sSVy9dlh

Inference APIs are filling sessions with encrypted reasoning, hidden search results, opaque compaction, and encrypted subagent messages. A growing form of lock-in.

15s
Read More
Automation Without Understanding
🖼️
0

Automation Without Understanding

Hacker News·3 months ago
#m3SMlCX8

Two developments are unfolding at once: artificial intelligence systems have begun to produce genuine research-level mathematics, and the United States is weakening the pipeline that produces humans capable of understanding what such systems are doing.…

15s
Read More
Prompt Injection as Role Confusion
🖼️
231

Prompt Injection as Role Confusion

Hacker News·4 months ago
#l5hU3EhN

LLMs can't tell who's speaking. We show they identify roles by writing style, not tags, and exploit this with CoT Forgery, injecting fake reasoning that models mistake for their own thoughts.

15s
Read More
VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models
🖼️
396

VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models

Hacker News·4 months ago
#pyu4kVuV

This technical report introduces VibeThinker-3B, a compact dense model with 3B parameters developed to investigate how far verifiable reasoning can be pushed within a strictly small-model regime.…

15s
Read More