Menu

#Models

1457 posts

4 points
Feed·
20 of 1457 posts
The Unbearable Cheapness of Open Weight Models – James O'Claire
📰
201

The Unbearable Cheapness of Open Weight Models – James O'Claire

Hacker News·3 months ago
#2oXAMeZO

Today I was setting up Hermes to see how it does with web research. I chose DeepSeek V4 because I know it is cheap, but seeing it’s pricing next to Anthropic and OpenAI ‘frontier’ models is crazy.…

15s
Read More
How to Choose the Right AI Model for Your Agent (2026 Decision Guide)
🖼️
1

How to Choose the Right AI Model for Your Agent (2026 Decision Guide)

DEV Community·Joaki·5 months ago
#qwxbrQ0z
#ai#agents#productivity#pick#gemini#model

Five years ago picking an AI model was a one-line decision. In 2026 there are six tied flagships, a dozen capable mid-tier models, and the wrong pick will either bankrupt you on tokens or cripple your

15s
Read More
What Goes Around Comes Around: A New Model Every Month and a Half
🖼️
1

What Goes Around Comes Around: A New Model Every Month and a Half

DEV Community·guanjiawei·5 months ago
#X77sj1oD
#ai#models#openai#claude#opus#model

GPT-5.5 dropped today, just a month and a half after 5.4. Opus 4.7 came out last week, just two months after 4.6. What's even funnier is that GPT finally started talking like a human, while Opus stopped doing so.…

15s
Read More
M5 Ultra Mac Studio Review: The Dream Mac for Local AI Agents
🖼️
0

M5 Ultra Mac Studio Review: The Dream Mac for Local AI Agents

Hacker News·16 days ago
#jDabESZR
#macstories#ultra#local#model#models#studio

For the past few days, I’ve been testing the (currently) top-of-the-line M5 Ultra Mac Studio with 256 GB of RAM. I’ll cut to the chase: the M5 Ultra Mac Studio is a dream machine for local AI agents.…

15s
Read More
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
🖼️
0

Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

Hacker News·22 days ago
#HCwA2Z5z
#gemini#live#google#voice#models#photo

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet, built for natural conversation.

15s
Read More
Why we built Pion | Andon Labs
🖼️
497

Why we built Pion | Andon Labs

Hacker News·23 days ago
#e0oXy0dk

Pion is the platform we built to run our autonomous businesses. Today we are opening it up so many more people can experiment with autonomous businesses, and here is why.

15s
Read More
On-device intelligence for every product
🖼️
0

On-device intelligence for every product

Hacker News·28 days ago
#bTROLJ2f
#desertant#model#models#device#every#article

We're building a frontier lab for on-device AI. Small, specialized models for audio, vision, and text, faster than the cloud and free.

15s
Read More
Introducing K2 Horizon: Frontier Performance, Radically Open
🖼️
335

Introducing K2 Horizon: Frontier Performance, Radically Open

Hacker News·about 1 month ago
#HfMIgQG8
#ifm#training#horizon#model#models#photo

Explore K2 Horizon, IFM’s open-source fleet of six frontier AI models for reasoning, coding, agentic workflows, edge devices, and enterprise deployment.

15s
Read More
Software Engineering fundamentals matter more than ever
📰
0

Software Engineering fundamentals matter more than ever

Hacker News·about 2 months ago
#8j2fKm3U

The manifestation of my imposter syndrome, for me and today, is what does it mean to be a software engineer. There’s a lot more noise than signal on the Internet about agentic engineering, wh…

15s
Read More
More models, more choice: Comparing 11 different AI models
🖼️
220

More models, more choice: Comparing 11 different AI models

Hacker News·about 2 months ago
#aUBhJxpg
#netlify#credits#model#models#photo#article

Netlify now runs any OpenRouter model, including Kimi K3, GLM 5.2 and DeepSeek V4. We tested 11 of them on the same build prompt to see how they differ.

15s
Read More
Emergent Introspective Awareness in Large Language Models
🖼️
72

Emergent Introspective Awareness in Large Language Models

Hacker News·about 2 months ago
#oRGvJBjd

We investigate whether large language models can introspect on their internal states. It is difficult to answer this question through conversation alone, as genuine introspection cannot be distinguished from confabulations.…

15s
Read More
📰
23

Ferrox: Building a Rust Inference Engine That Matches llama.cpp

Hacker News·2 months ago
#XqbMNFIk
#fratepietro#llama#ferrox#metal#gguf#models

Why I built Ferrox, a pure-Rust GGUF inference engine, and what it took to match llama.cpp’s performance on real models — with receipts, not vibes.

15s
Read More
Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence
🖼️
175

Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence

Hacker News·2 months ago
#r9qoYdCX

Both the general public and academic communities have raised concerns about sycophancy, the phenomenon of artificial intelligence (AI) excessively agreeing with or flattering users.…

15s
Read More