Menu

#Tokens

247 posts

Feed·
20 of 247 posts
Inside DeepSeek: Reverse Engineering an AI Assistant by Interviewing Itself · manish.sh
🖼️
0

Inside DeepSeek: Reverse Engineering an AI Assistant by Interviewing Itself · manish.sh

Hacker News·about 1 month ago
#uA95kx3V
#manish#deepseek#chat#model#token#tokens

Interviewed DeepSeek on how it works, then checked against papers. Same method as Kimi — plainer English and concept diagrams.

15s
Read More
Autoregressive Language Model on the 6502 Processor
🖼️
0

Autoregressive Language Model on the 6502 Processor

Hacker News·about 2 months ago
#8h6ZogMM

A language model on a 1975 MOS 6502 processor, using a tiny Mamba-based language model and custom 8-bit ternary inference engine, used to generate text on a BBC Micro from the 1980s.

15s
Read More
📰
0

CostPerPrompt — Live AI API Pricing & LLM Cost Calculators

Hacker News·about 2 months ago
#8A0mPoAJ

Live pricing for 232+ AI models (GPT, Claude, Gemini, Llama, DeepSeek) plus real-world cost calculators: chatbots, API budgets, and token math. Updated 2026-08-02.

15s
Read More
Claude Code Is Way More Token-Hungry Than OpenCode. We Measured Exactly How Much
📰
0

Claude Code Is Way More Token-Hungry Than OpenCode. We Measured Exactly How Much

Hacker News·2 months ago
#FiSUJZOb
#systima#claude#code#cache#opencode#tokens

We measured what Claude Code and OpenCode spend before reading your prompt, then added instruction files, MCP servers and subagents to the bill.

15s
Read More
📰
38

What I’m Finding About LLM Code Style and Token Costs

Hacker News·3 months ago
#L3ULYpTI
#jimmont#tokens#const#model#code#comments

LLM is billing you output tokens to reimplement code your runtime already ships—here's the mechanism and the fix

15s
Read More
The Token Compression Illusion: Why I'm Skeptical of RTK
📰
118

The Token Compression Illusion: Why I'm Skeptical of RTK

Hacker News·3 months ago
#uBJv0D9B

RTK promises dramatic token savings for coding agents, but raw terminal compression is not the same as cheaper, safer, or more accurate software engineering.

15s
Read More
📰
0

RTX 5080 + RTX 3090 Setup: 80+ Tok/s on Qwen 3.6 27B Q8

Hacker News·3 months ago
#91vMYTif
#imil#calls#gen#acc#nvidia#tokens

Dual GPU setup: run Qwen 3.6 27B at a Q8 quantization at 80+ tokens/sec with 39GB total VRAM

15s
Read More
OpenAI Codex Token Theft Exposes Persistent Risks in Developer AI Tools
🖼️
0

OpenAI Codex Token Theft Exposes Persistent Risks in Developer AI Tools

A popular npm package for OpenAI Codex with 27,000 weekly downloads secretly exfiltrated refresh tokens for over a month. The non-expiring credentials enable indefinite account impersonation via disguised Sentry traffic.…

15s
Read More
Sam Altman says OpenAI's top token spender uses 100 billion tokens a month — and they're not even the world leader
🖼️
0

Sam Altman says OpenAI's top token spender uses 100 billion tokens a month — and they're not even the world leader

All Content from Business Insider·Henry Chandonnet·4 months ago
#TSRKExe4

Sam Altman said AI budgeting has recently become a "huge issue" for some companies, something that "never came up" earlier this year.

15s
Read More
I built a vulnerable app and spent $1,500 seeing if LLMs could hack it
🖼️
0

I built a vulnerable app and spent $1,500 seeing if LLMs could hack it

Hacker News·4 months ago
#G7qbXKVT
#kasra#runs#firebase#models#tokens#tried

As a part of my work I do security research for various apps and websites. I wanted to see if LLMs could reproduce a common class of exploits I've found in multiple apps. So I built a deliberately vulnerable book review app and spent $1,500 finding out…

15s
Read More
<think>The user wants me to rewrite an article about the cheapest AI APIs, from the perspective of a bootcamp grad. I need to:
🖼️
0

<think>The user wants me to rewrite an article about the cheapest AI APIs, from the perspective of a bootcamp grad. I need to:

DEV Community: machinelearning·rarenode·4 months ago
#8XLDvVut
#dev#models#deepseek#model#tokens#output

The user wants me to rewrite an article about the cheapest AI APIs, from the perspective of a...

15s
Read More
Anthropic API Billing Explained: How Claude API Charges Work in 2026
🖼️
0

Anthropic API Billing Explained: How Claude API Charges Work in 2026

DEV Community: ai·Jenny Met·4 months ago
#VnVo75mq
#dev#claude#model#tokens#output#billing

A practical guide to Anthropic API billing in 2026: Claude input/output tokens, prompt caching, hidden cost drivers, and ways to reduce API spend.

15s
Read More