Menu

#Token

415 posts

Feed·
20 of 415 posts
GitHub - argonautlabsai/deltafin: ARGODRIVE Deltafin: Kimi K3 (2.8T MoE) streamed from SSDs on Apple Silicon — fork of gavamedia/deltafin with the ARGODRIVE storage work and benchmark package
🖼️
0

GitHub - argonautlabsai/deltafin: ARGODRIVE Deltafin: Kimi K3 (2.8T MoE) streamed from SSDs on Apple Silicon — fork of gavamedia/deltafin with the ARGODRIVE storage work and benchmark package

Hacker News·28 days ago
#GbimiaTR
#github#deltafin#token#full#release#prompt

ARGODRIVE Deltafin: Kimi K3 (2.8T MoE) streamed from SSDs on Apple Silicon — fork of gavamedia/deltafin with the ARGODRIVE storage work and benchmark package - argonautlabsai/deltafin

15s
Read More
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems
🖼️
79

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

Hacker News·about 1 month ago
#GhvmCEv5
#arxiv#context#cost#token#view#production

Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning context: conversation histories, large prompts, large tool definitions, and ballooning tool outputs.…

15s
Read More
GitHub - zachahn/vomit: Clean up Claude 5's token vomit with a separate LLM. Save your tokens, Claude 5 is hopeless
📰
0

GitHub - zachahn/vomit: Clean up Claude 5's token vomit with a separate LLM. Save your tokens, Claude 5 is hopeless

Hacker News·about 2 months ago
#MAXWBPjb
#github#claude#llm#local#token#article

Clean up Claude 5's token vomit with a separate LLM. Save your tokens, Claude 5 is hopeless - zachahn/vomit

15s
Read More
Inside DeepSeek: Reverse Engineering an AI Assistant by Interviewing Itself · manish.sh
🖼️
0

Inside DeepSeek: Reverse Engineering an AI Assistant by Interviewing Itself · manish.sh

Hacker News·about 2 months ago
#uA95kx3V
#manish#deepseek#chat#model#token#tokens

Interviewed DeepSeek on how it works, then checked against papers. Same method as Kimi — plainer English and concept diagrams.

15s
Read More
Needle 2 - The 14 MB Agentic LLM for Tiny Devices | Cactus
🖼️
527

Needle 2 - The 14 MB Agentic LLM for Tiny Devices | Cactus

Hacker News·about 2 months ago
#0kJW6fQw

An open 45M-parameter model for tool calling, device use, and structured extraction. Needle 2 runs as a 14 MB binary in 28 MB of session RAM.

15s
Read More
Petri Nets as a Music Sequencer
🖼️
76

Petri Nets as a Music Sequencer

Hacker News·2 months ago
#Ag0Rwao1
#token#circles#ring#track#beats#article

A music sequencer built entirely on Petri nets — token rings become drum machines, Euclidean rhythms fall out of the topology, and polyrhythm comes free.

15s
Read More
Keyv and friends compromised in active Shai-Hulud supply chain attack
🖼️
250

Keyv and friends compromised in active Shai-Hulud supply chain attack

Hacker News·2 months ago
#XsFRuANf
#aikido#github#month#payload#token#files

Mini Shai-Hulud malware was injected into keyv and eight related npm packages on August 4, 2026 after an attacker compromised the maintainer's GitHub account

15s
Read More
Autoregressive Language Model on the 6502 Processor
🖼️
0

Autoregressive Language Model on the 6502 Processor

Hacker News·2 months ago
#8h6ZogMM

A language model on a 1975 MOS 6502 processor, using a tiny Mamba-based language model and custom 8-bit ternary inference engine, used to generate text on a BBC Micro from the 1980s.

15s
Read More
GitHub - sqliteai/waste: Run the full 2.78-trillion-parameter Kimi K3 model beyond available RAM by streaming activated weights directly from NVMe. A dependency-free, embeddable C inference engine.
📰
341

GitHub - sqliteai/waste: Run the full 2.78-trillion-parameter Kimi K3 model beyond available RAM by streaming activated weights directly from NVMe. A dependency-free, embeddable C inference engine.

Hacker News·2 months ago
#i9E3YRcs
#github#container#waste#cache#token#engine

Run the full 2.78-trillion-parameter Kimi K3 model beyond available RAM by streaming activated weights directly from NVMe. A dependency-free, embeddable C inference engine. - sqliteai/waste

15s
Read More
aistack - How many devs can you fit on a GPU?
📰
0

aistack - How many devs can you fit on a GPU?

Hacker News·2 months ago
#FVRCoQnC
#aistack#model#hardware#token#cost#tasks

aistack is an open benchmarking initiative by imec, generating solid datapoints where real AI workloads meet real systems and silicon.

15s
Read More
My security camera shipped a GitHub admin token in its login page
📰
0

My security camera shipped a GitHub admin token in its login page

Hacker News·2 months ago
#u0rQr9ml
#hhh#hanwha#token#cameras#github#article

i dissected some firmware for a Hanwha Wisenet XNP-9300RW and found that it had github admin tokens in it, also miltech is weird

15s
Read More
Qwen3.8 is launching and going open-weight soon!🌐

With a massive 2.4T parameters, this model is continuously evolving. We believe it’s one of the most powerful model available today, compatible to leading frontier AI models , second only to Fable 5.

You don't have to wait to https://t.co/JS3ID73IYS
🖼️
0

Qwen3.8 is launching and going open-weight soon!🌐 With a massive 2.4T parameters, this model is continuously evolving. We believe it’s one of the most powerful model available today, compatible to leading frontier AI models , second only to Fable 5. You don't have to wait to https://t.co/JS3ID73IYS

Hacker News·3 months ago
#6hPsiVYT
#twitter#token#qwen3#model#wait#plan

Qwen3.8 is launching and going open-weight soon!🌐 With a massive 2.4T parameters, this model is continuously evolving. We believe it’s one of the most powerful model available today, compatible to leading frontier AI models , second only to Fable 5.…

15s
Read More
Token-Savior's 5 Hidden Uses: The MCP Server That Cuts Your AI Coding Costs by 80%
🖼️
0

Token-Savior's 5 Hidden Uses: The MCP Server That Cuts Your AI Coding Costs by 80%

DEV Community·韩·4 months ago
#lXPpUo1r
#dev#token#claude#savior#code#hidden

You probably use Claude Code, Cursor, or Windsurf every day — but if you're not running an MCP server...

15s
Read More
Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering
🖼️
0

Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering

Hacker News·4 months ago
#vC8iQUyC

LLM-based Multi-Agent (LLM-MA) systems are increasingly applied to automate complex software engineering tasks such as requirements engineering, code generation, and testing.…

15s
Read More