Menu

#Llm

736 posts

Feed·
20 of 736 posts
Language Models for Text Classification: From Bag-of-Words to Jev
🖼️
213

Language Models for Text Classification: From Bag-of-Words to Jev

Hacker News·8 days ago
#P5jggCFw
#magazine#jev#llm#classifier#article#ama

A Visual Guide to RNNs, CNNs, Transformers, and Calibration, with Hands-On Experiments on Accuracy and Efficiency

15s
Read More
How to win a beer with high-dimensional statistics
🖼️
81

How to win a beer with high-dimensional statistics

Hacker News·9 days ago
#CXU6baEh

My longtime labmate-turned-student/friend1 Dhruva Karkada recently wrote a sick paper on data statistics which deservedly went viral on Twitter, in part beca...

15s
Read More
📰
283

MicroLLM lab — tiny LLMs, Q4, in your browser

stateofutopia.com·9 days ago
#tZWWiq90

Run Q4 small language models entirely in the browser with WebGPU. Chat, benchmark, and compare PetitGPT, SmolLM2, and friends on-device.

15s
Read More
How to keep enjoying programming in a world of LLMs
📰
352

How to keep enjoying programming in a world of LLMs

Hacker News·11 days ago
#iPrR5Rx5

Are you steering towards AI burnout? Afraid of loosing your job to someone with little programming skills, no aspirations to quality, and a huge Claude account? Disappointed about the code quality in your projects, or worse in “your” own code?…

15s
Read More
Turn GLM-5.3-Flash into a Jev-like System One model
🖼️
138

Turn GLM-5.3-Flash into a Jev-like System One model

Hacker News·11 days ago
#vG6hnZy6

Typed decisions with a probability for every option, from GLM-5.3-Flash in a single forward pass: about as accurate as Jev, faster from Europe, and protected by confidential computing. With a playground and a benchmark on 29 datasets.

15s
Read More
DeepSeek Elastic Compute (DSec): A Sandbox Infrastructure for Effective Agentic Training at Scale
🖼️
323

DeepSeek Elastic Compute (DSec): A Sandbox Infrastructure for Effective Agentic Training at Scale

Hacker News·11 days ago
#noJZrAl4
#arxiv#deepseek#llm#sandbox#dsec#photo

Large-scale agentic training and evaluation with large language models (LLMs) rely on isolated, stateful execution environments in which models inspect repositories, invoke tools, execute commands, and interact with task-specific services.…

15s
Read More
A Jev-like wrapper for LLMs, including vision models
🖼️
157

A Jev-like wrapper for LLMs, including vision models

Hacker News·12 days ago
#SgosVDAQ
#allanrbo#llm#logprobs#jev#opencv#article

I was intrigued by Jev and the self-hostable projects appearing around it, such as OpenJev and SemIf . Reading about them introduced me ...

15s
Read More
SlopShape: Identifying AI-Generated Commercial Web Content
🖼️
0

SlopShape: Identifying AI-Generated Commercial Web Content

Hacker News·15 days ago
#c1dvrYhN
#arxiv#llm#photo#englishlanguage

Word-level detectors identify unedited AI-generated text almost perfectly, but the literature documents their brittleness under rewording, and a word-level score neither characterizes a text nor identifies which AI model wrote it.…

15s
Read More
Will OpenAI Eat Jev's Lunch?
🖼️
0

Will OpenAI Eat Jev's Lunch?

Hacker News·15 days ago
#d5fEbEso
#arcturus-labs#openai#jev#llm#typesafe#ai

TypeSafe's Jev is a genuine breakthrough – snap judgments with calibrated probabilities instead of generated text. My bet is OpenAI is already figuring out how to copy it, and then embed it inside its own models where Jev can't follow.

15s
Read More
📰
0

</> htmx ~ Markdown in /src

htmx.org·15 days ago
#Lt3GwsSV

In this essay, Carson Gross argues that Markdown is now an important source artifact for software systems built with LLMs, and that, following the principle of locality, it should live in /src, alongside the code it produces.

15s
Read More