Menu

#Arxiv

79 posts

1 point
Feed·
20 of 79 posts
mold: A Massively Parallel Linker
🖼️
187

mold: A Massively Parallel Linker

Hacker News·about 1 month ago
#jxGVgApS

Linking is a critical step in the software build process that combines compiled object files into a single executable or shared library. Despite decades of engineering effort, link times remain a significant bottleneck in the edit-compile-debug cycle,…

15s
Read More
Thinking Fast and Slow in AI: the Role of Metacognition
🖼️
177

Thinking Fast and Slow in AI: the Role of Metacognition

Hacker News·9 days ago
#jbAySySO

AI systems have seen dramatic advancement in recent years, bringing many applications that pervade our everyday life. However, we are still mostly seeing instances of narrow AI: many of these recent developments are typically focused on a very limited set…

15s
Read More
DeepSeek Elastic Compute (DSec): A Sandbox Infrastructure for Effective Agentic Training at Scale
🖼️
323

DeepSeek Elastic Compute (DSec): A Sandbox Infrastructure for Effective Agentic Training at Scale

Hacker News·10 days ago
#noJZrAl4
#arxiv#deepseek#llm#sandbox#dsec#photo

Large-scale agentic training and evaluation with large language models (LLMs) rely on isolated, stateful execution environments in which models inspect repositories, invoke tools, execute commands, and interact with task-specific services.…

15s
Read More
arXiv receives Multiyear Philanthropic Commitments to Support Its Launch as an Independent Nonprofit
🖼️
0

arXiv receives Multiyear Philanthropic Commitments to Support Its Launch as an Independent Nonprofit

Hacker News·13 days ago
#CZGk84Jy

arXiv is pleased to announce that three leading philanthropic organizations have provided multimillion-dollar support for arXiv. These new multiyear…

15s
Read More
SlopShape: Identifying AI-Generated Commercial Web Content
🖼️
0

SlopShape: Identifying AI-Generated Commercial Web Content

Hacker News·14 days ago
#c1dvrYhN
#arxiv#llm#photo#englishlanguage

Word-level detectors identify unedited AI-generated text almost perfectly, but the literature documents their brittleness under rewording, and a word-level score neither characterizes a text nor identifies which AI model wrote it.…

15s
Read More
Semantics for 2D Rasterization
🖼️
143

Semantics for 2D Rasterization

Hacker News·17 days ago
#RrInDhXK

Rasterization is the process of determining the color of every pixel drawn by an application. Powerful rasterization libraries like Skia, CoreGraphics, and Direct2D put exceptional effort into drawing, blending, and rendering efficiently.…

15s
Read More
The Implications of Linguistic Illegibility for LLM Security
🖼️
0

The Implications of Linguistic Illegibility for LLM Security

Hacker News·18 days ago
#VMdJEQEX

LLMs are trained to generate natural language. However, various strands of evidence indicate that an LLM's externalized linguistic outputs and mechanistically-extracted linguistic features can be an unreliable lens for understanding internal model…

15s
Read More
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
🖼️
0

Cache-to-Cache: Direct Semantic Communication Between Large Language Models

Hacker News·18 days ago
#P5iyALTD

Multi-LLM systems harness the complementary strengths of diverse Large Language Models, achieving performance and efficiency gains that are not attainable by a single model.…

15s
Read More
An Empirical Study of Harness Design for Coding Agents
🖼️
0

An Empirical Study of Harness Design for Coding Agents

Hacker News·18 days ago
#9ELYdVBy

Coding harnesses shape how autonomous coding agents translate model capabilities into long-horizon software-engineering performance, yet existing work typically evaluates harnesses as monolithic systems, leaving the effectiveness of individual components…

15s
Read More
Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data
🖼️
0

Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data

Hacker News·19 days ago
#6i69bOIu

The scaling laws hold that a language model grows more capable with more parameters and more training data, and Mixture-of-Experts (MoE) architectures have ridden these laws to remarkable results, activating only a fraction of an enormous stored parameter…

15s
Read More
Dream-RSI: Recursive Self-Improvement through Evolving Worlds
🖼️
0

Dream-RSI: Recursive Self-Improvement through Evolving Worlds

Hacker News·20 days ago
#kF41kvJy

Recursive self-improvement is becoming increasingly vital for autonomous AI agents, where progress hinges on discovering high-value solutions across complex domains.…

15s
Read More
Breaking the 1.58-bit Barrier for Ternary LLMs
🖼️
0

Breaking the 1.58-bit Barrier for Ternary LLMs

Hacker News·20 days ago
#2bgCT72I

Ternary Large Language Models (LLM) store every weight as one of three symbols $\{-1,0,+1\}$, so the cost of a ternary model is conventionally referenced to the information-theoretic $\log_2 3 \approx 1.585$ bits per weight.…

15s
Read More
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
🖼️
0

The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation

Hacker News·23 days ago
#s2hr6cLP
#arxiv#miles#brundage#view#landscape#threats

This report surveys the landscape of potential security threats from malicious uses of AI, and proposes ways to better forecast, prevent, and mitigate these threats.…

15s
Read More
Will there be a 7G?
🖼️
0

Will there be a 7G?

Hacker News·24 days ago
#yRrhKWzx
#arxiv#6g#7g#wireless#standards#photo

The transition from 5G to 6G is becoming concrete: the ITU-R IMT-2030 framework has established the high-level vision and capability set for 6G, while 3GPP Release 21 has defined the path toward the first 6G specifications.…

15s
Read More
Trusting-Trust Attack against an Entire Linux Distribution through Binary Manipulation
🖼️
0

Trusting-Trust Attack against an Entire Linux Distribution through Binary Manipulation

Hacker News·29 days ago
#K4DnZKkt
#arxiv#attack#strip#view#trusting#trust

Ken Thompson's trusting-trust attack, in which a compromised compiler backdoors the programs it builds and reproduces the backdoor in subsequent rebuilds of itself, is widely regarded as a threat specific to compilers. We show that it is not.…

15s
Read More
Higher multipoles of the cow
🖼️
0

Higher multipoles of the cow

Hacker News·about 1 month ago
#2NEocNeG

The spherical cow approximation is widely used in the literature, but is rarely justified. Here, I propose several schemes for extending the spherical cow approximation to a full multipole expansion, in which the spherical cow is simply the first term.…

15s
Read More
Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment
🖼️
117

Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment

Hacker News·about 1 month ago
#tzHPXVYg

We study autonomous mathematical discovery in the Station, an open-world multi-agent environment in which AI agents from different model families pursue a shared research goal without a central coordinator or scripted pipeline.…

15s
Read More
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems
🖼️
79

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

Hacker News·about 1 month ago
#GhvmCEv5
#arxiv#context#cost#token#view#production

Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning context: conversation histories, large prompts, large tool definitions, and ballooning tool outputs.…

15s
Read More
Black hole singularity is a surface not a point
🖼️
298

Black hole singularity is a surface not a point

Hacker News·about 1 month ago
#iZPLxxQk
#arxiv#black#hole#singularity#surface#quantum

It is widely repeated in the popular literature and elsewhere that the singularity at the center of a black hole is a point. It is not true. Two observers who free-fall into a spherical black hole along two different angular trajectories at the same time…

15s
Read More
Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!
🖼️
0

Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!

Hacker News·about 2 months ago
#Zhv3u9jl

Intermediate token generation (ITG), where a model produces output before the solution, has become a standard method to improve the performance of language models on reasoning tasks.…

15s
Read More