AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,157 results
5 May 2026

// HeavySkill // One of the cleaner takes on agentic harness design I've read. They argue that what actually drives agent harness performanc…

Model ReleasesDGX agent

// HeavySkill // One of the cleaner takes on agentic harness design I've read. They argue that what actually drives agent harness performance is not the orchestration code. It's a single inner skill:

How Prompts Move Language Model Behavior: Frames, Salience, and Construal as Semantic Control

ResearchDGX agent

arXiv:2512.12688v3 Announce Type: replace-cross Abstract: Prompt engineering is widely used to shape large language model behavior, yet it is often treated as a practical heuristic rather than as a fo

HumanSplatHMR: Closing the Loop Between Human Mesh Recovery and Gaussian Splatting Avatar

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.02784v1 Announce Type: new Abstract: Accurately recovering human pose and appearance from video is an essential component of scene reconstruction, with applications to motion capture, motio

Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

Model ReleasesDGX agent

arXiv:2605.02801v1 Announce Type: new Abstract: As large language model (LLM) agents evolve from isolated tool users into coordinated teams, reinforcement learning (RL) must optimize not only individu

Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor

ResearchDGX agent

arXiv:2208.00335v4 Announce Type: replace Abstract: Rule extraction is a central problem in interpretable machine learning because it seeks to convert opaque predictive behavior into human-readable sy

TIJERE: A Novel Threat Intelligence Joint Extraction Model Based on Analyst Expert Knowledge

ApplicationsDGX agent

arXiv:2605.02041v1 Announce Type: new Abstract: The extraction of entities and relationships from threat intelligence reports into structured formats, such as cybersecurity knowledge graphs, is essent

Unbox Responsible GeoAI: Navigating Climate Extreme and Disaster Mapping

ResearchDGX agent

arXiv:2605.00315v1 Announce Type: cross Abstract: As climate extreme and disaster events become more frequent and intense, Geospatial Artificial Intelligence (GeoAI) has emerged as a transformative ap

4 May 2026

Can Coding Agents Reproduce Findings in Computational Materials Science?

Model ReleasesDGX agent

arXiv:2605.00803v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous coding agents and have achieved remarkably strong performance on software engineering be

Efficient Mutation Testing of Quantum Machine Learning Models

TutorialsDGX agent

arXiv:2605.00107v1 Announce Type: cross Abstract: Quantum machine learning integrates the strengths of quantum computing and machine learning, enabling models to learn complex features using fewer par

2 May 2026

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compa…

SafetyDGX agent

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compared to largely speculative arguments about “instrumental con

Claude Opus 4.7 just implemented an AlphaZero-style self-play pipeline from scratch. It did this on consumer hardware in three hours, then b…

Model ReleasesDGX agent

Claude Opus 4.7 just implemented an AlphaZero-style self-play pipeline from scratch. It did this on consumer hardware in three hours, then beat the Pascal Pons solver 7 of 8 as first-mover on Connect

1 May 2026

AI Inference as Relocatable Electricity Demand: A Latency-Constrained Energy-Geography Framework

ApplicationsDGX agent

arXiv:2604.27855v1 Announce Type: cross Abstract: AI inference is becoming a persistent and geographically distributed source of electricity demand. Unlike many traditional electrical loads, inference

Evaluating Epistemic Guardrails in AI Reading Assistants: A Behavioral Audit of a Minimal Prototype

ResearchDGX agent

arXiv:2604.27275v1 Announce Type: cross Abstract: Large language model (LLM) reading assistants are increasingly used in settings that require interpretation rather than simple retrieval. In these con

Mixed Precision Training of Neural ODEs

ResearchDGX agent

arXiv:2510.23498v2 Announce Type: replace-cross Abstract: Exploiting low-precision computations has become a standard strategy in deep learning to address the growing computational costs imposed by ev

Remaining Useful Life Estimation for Turbofan Engines: A Comparative Study of Classical, CNN, and LSTM Approaches

ResearchDGX agent

arXiv:2604.27234v1 Announce Type: new Abstract: Remaining Useful Life (RUL) estimation is a critical component of Prognostics and Health Management (PHM), enabling proactive maintenance scheduling and

Semantic Variational Bayes Based on Semantic Information G Theory for Solving Latent Variables

Model ReleasesDGX agent

arXiv:2408.13122v2 Announce Type: replace-cross Abstract: The Variational Bayesian method (VB) is used to solve the probability distributions of latent variables with the minimum free energy criterion

Simple Self-Conditioning Adaptation for Masked Diffusion Models

ResearchDGX agent

arXiv:2604.26985v1 Announce Type: cross Abstract: Masked diffusion models (MDMs) generate discrete sequences by iterative denoising under an absorbing masking process. In standard masked diffusion, if

Theory Under Construction: Orchestrating Language Models for Research Software Where the Specification Evolves

Model ReleasesDGX agent

arXiv:2604.27209v1 Announce Type: cross Abstract: Large language models can now generate substantial code and draft research text, but research-software projects require more than either artifact alon

30 Apr 2026

Data Balancing Strategies: A Systematic Survey of Resampling and Augmentation Methods

ResearchDGX agent

arXiv:2505.13518v2 Announce Type: replace-cross Abstract: Imbalanced datasets, where one class significantly outnumbers others, remain a persistent challenge in machine learning, often biasing predict

LATTICE: Evaluating Decision Support Utility of Crypto Agents

Model ReleasesDGX agent

arXiv:2604.26235v1 Announce Type: cross Abstract: We introduce LATTICE, a benchmark for evaluating the decision support utility of crypto agents in realistic user-facing scenarios. Prior crypto agent

29 Apr 2026

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving

AgentsDGX agent

arXiv:2408.16322v4 Announce Type: replace Abstract: Current research in semantic bird's-eye view segmentation for autonomous driving focuses solely on optimizing neural network models using a single d

Is the Modality Gap a Bug or a Feature? A Robustness Perspective

ApplicationsDGX agent

arXiv:2603.29080v2 Announce Type: replace Abstract: Many modern multi-modal models (e.g. CLIP) seek an embedding space in which the two modalities are aligned. Somewhat surprisingly, almost all existi

28 Apr 2026

ACL ARR March 2026 Cycle [D]

ResearchDGX agent

The ACL Rolling Review (ARR) operates on a two-monthly review cycle , and the March 2026 cycle refers to one of these periodic submission and peer review rounds for computational linguistics research.

Agentic Witnessing: Pragmatic and Scalable TEE-Enabled Privacy-Preserving Auditing

Model ReleasesDGX agent

arXiv:2604.24203v1 Announce Type: cross Abstract: Auditing the semantic properties of proprietary data creates a fundamental tension: verification requires transparent access, while proprietary rights

On the Convergence Theory of Pipeline Gradient-based Analog In-memory Training

ResearchDGX agent

arXiv:2410.15155v3 Announce Type: replace Abstract: Aiming to accelerate the training of large deep neural networks (DNN) in an energy-efficient way, analog in-memory computing (AIMC) emerges as a sol

PEPS: Positional Encoding Projected Sampling -- Extended

TutorialsDGX agent

arXiv:2604.24167v1 Announce Type: new Abstract: Implicit neural representations (INRs) are increasingly being used as tools to map coordinates to signals, encompassing applications from neural fields

Revisable by Design: A Theory of Streaming LLM Agent Execution

AgentsDGX agent

arXiv:2604.23283v1 Announce Type: new Abstract: Current LLM agents operate under an implicit but universal assumption: execution is a transaction -- the user submits a request, the agent works in isol

Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk

Model ReleasesDGX agent

arXiv:2604.24197v1 Announce Type: cross Abstract: Frontier image generation has moved from artistic synthesis toward synthetic visual evidence. Systems such as GPT Image 2, Nano Banana Pro, Nano Banan

Towards Automated Ontology Generation from Unstructured Text: A Multi-Agent LLM Approach

AgentsDGX agent

arXiv:2604.23090v1 Announce Type: new Abstract: Automatically generating formal ontologies from unstructured natural language remains a central challenge in knowledge engineering. While large language

27 Apr 2026

'AI should elevate your thinking, not replace it.' I don't disagree, but the issue is that current LLMs are not really trained to support th…

AgentsDGX agent

'AI should elevate your thinking, not replace it.' I don't disagree, but the issue is that current LLMs are not really trained to support that out of the box. I've solved this by building my own agent

Associativity-Peakiness Metric for Contingency Tables

ResearchDGX agent

arXiv:2604.22655v1 Announce Type: new Abstract: For the use case of comparing the performance of clustering algorithms whose output is a contingency table, a single performance metric for contingency

BLAST: Benchmarking LLMs with ASP-based Structured Testing

ResearchDGX agent

arXiv:2604.22306v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable performance across a broad spectrum of tasks, including natural language understanding, dial

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model

SafetyDGX agent

arXiv:2604.22152v1 Announce Type: new Abstract: Evaluating robotics policies across thousands of environments and thousands of tasks is infeasible with existing approaches. This motivates the need for

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we exp…

AgentsDGX agent

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we explored what happens if we train an AI to do that job instead.

Generating Synthetic Malware Samples Using Generative AI

ResearchDGX agent

arXiv:2604.22084v1 Announce Type: new Abstract: Malware attacks have a significant negative impact on organizations of varied scales in the field of cybersecurity. Recently, malware researchers have i

Looking forward to the work coming out of @IneffableLabs 🇬🇧 Largest EU/UK raise ever 😮 Beat the previous UK largest seed raise ($101m Sta…

IndustryDGX agent

Looking forward to the work coming out of @IneffableLabs 🇬🇧 Largest EU/UK raise ever 😮 Beat the previous UK largest seed raise (101m Stability AI!) by 999m 🚀 We have seen amazing things on our self-le

QDTraj: Exploration of Diverse Trajectory Primitives for Articulated Objects Robotic Manipulation

AgentsDGX agent

arXiv:2604.22551v1 Announce Type: cross Abstract: Thanks to the latest advances in learning and robotics, domestic robots are beginning to enter homes, aiming to execute household chores autonomously.

Sound Agentic Science Requires Adversarial Experiments

AgentsDGX agent

arXiv:2604.22080v1 Announce Type: new Abstract: LLM-based agents are rapidly being adopted for scientific data analysis, automating tasks once limited by human time and expertise. This capability is o

Teaching an Agent to Sketch One Part at a Time

AgentsDGX agent

arXiv:2603.19500v2 Announce Type: replace Abstract: We develop a method for producing vector sketches one part at a time. To do this, we train a multi-modal language model-based agent using a novel mu

Very cool analysis of the submissions to a major management journal that shows how much the system of science, built for humans, is under st…

ApplicationsDGX agent

Very cool analysis of the submissions to a major management journal that shows how much the system of science, built for humans, is under strain as a result of AI. AI can be used to do better science

26 Apr 2026

Here is a very common problem when building complex agents. Long-horizon agents (in particular) fail in two ways: the decision-maker can't d…

SafetyDGX agent

Here is a very common problem when building complex agents. Long-horizon agents (in particular) fail in two ways: the decision-maker can't decompose well, or the skill library goes stale. This new res

👀 The most interesting people in the most exotic places can occasionally be persuaded to talk about NFTs, digital collectibles, and even @C…

IndustryDGX agent

👀 The most interesting people in the most exotic places can occasionally be persuaded to talk about NFTs, digital collectibles, and even @CandyDigital! Try it sometime just in case it catches on… 🔥🚀📈

25 Apr 2026

Alpha Eval: Agents Making Evals as a Multi-Player Game This diagram and blurb is largely a research riff with Claude on building data genera…

Model ReleasesDGX agent

Alpha Eval: Agents Making Evals as a Multi-Player Game This diagram and blurb is largely a research riff with Claude on building data generation systems to get closer to the holy grail of self-improvi

Bet this happens with Navier Stokes and it’s going to be something not even related to PDEs that solves it

IndustryDGX agent

Bet this happens with Navier Stokes and it’s going to be something not even related to PDEs that solves it 23 years old with no advanced mathematics training solves Erdős problem with ChatGPT Pro. 'Wh

24 Apr 2026

DeepSeek V4 - almost on the frontier, a fraction of the price

Model ReleasesDGX agent

Chinese AI lab DeepSeek's last model release was V3.2 (and V3.2 Speciale) last December. They just dropped the first of their hotly anticipated V4 series in the shape of two preview models, DeepSeek-V

Losing our Tail, Again: (Un)Natural Selection & Multilingual LLMs

ResearchDGX agent

arXiv:2507.03933v3 Announce Type: replace Abstract: Multilingual Large Language Models considerably changed how technologies influence language. While previous technologies could mediate or assist hum

PosterForest: Hierarchical Multi-Agent Collaboration for Scientific Poster Generation

AgentsDGX agent

arXiv:2508.21720v2 Announce Type: replace Abstract: Automating scientific poster generation requires hierarchical document understanding and coherent content-layout planning. Existing methods often re

RewardBench 2: Advancing Reward Model Evaluation

Model ReleasesDGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

Scaling of Gaussian Kolmogorov--Arnold Networks

Model ReleasesDGX agent

arXiv:2604.21174v1 Announce Type: cross Abstract: The Gaussian scale parameter (epsilon) is central to the behavior of Gaussian Kolmogorov--Arnold Networks (KANs), yet its role in deep edge-based arch

Verifying Machine Learning Interpretability Requirements through Provenance

ResearchDGX agent

arXiv:2604.21599v1 Announce Type: cross Abstract: Machine Learning (ML) Engineering is a growing field that necessitates an increase in the rigor of ML development. It draws many ideas from software e

23 Apr 2026

5B tokens in ml-intern in 48h 😅😅😅

IndustryDGX agent

5B tokens in ml-intern in 48h 😅😅😅 we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a great business strategy 🙈 So

Artifacts of Numerical Integration in Learning Dynamical Systems

AgentsDGX agent

arXiv:2507.14491v3 Announce Type: replace-cross Abstract: In many applications, one needs to learn a dynamical system from its solutions sampled at a finite number of time points. The learning problem

If you can’t stop small teams from using your API for distillation, then you’re definitely not stopping criminals, biohackers, or adversaria…

IndustryDGX agent

If you can’t stop small teams from using your API for distillation, then you’re definitely not stopping criminals, biohackers, or adversarial states from using AI through it. Those actors are much mor

Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem

SafetyDGX agent

arXiv:2604.20805v1 Announce Type: cross Abstract: The value alignment problem for artificial intelligence (AI) is often framed as a purely technical or normative challenge, sometimes focused on hypoth

RExBench: Can coding agents autonomously implement AI research extensions?

Model ReleasesDGX agent

arXiv:2506.22598v3 Announce Type: replace Abstract: Agents based on Large Language Models (LLMs) have shown promise for performing sophisticated software engineering tasks autonomously. In addition, t

Towards Explainable Federated Learning: Understanding the Impact of Differential Privacy

ResearchDGX agent

arXiv:2602.10100v2 Announce Type: replace Abstract: Data privacy and eXplainable Artificial Intelligence (XAI) are two important aspects for modern Machine Learning systems. To enhance data privacy, r

We still don't have a more precise value for 'Big G'

IndustryDGX agent

A decade-long precision measurement of the gravitational constant (G) at NIST yielded a value 0.0235% lower than a previous French experiment , but the result does not resolve the uncertainty in G, wh

22 Apr 2026

A couple of months ago, I announced that I was partway through implementing a simple, readable AlphaFold2 in pure PyTorch, inspired by @karp…

IndustryDGX agent

A couple of months ago, I announced that I was partway through implementing a simple, readable AlphaFold2 in pure PyTorch, inspired by @karpathy's minGPT. Today, I'm happy to share minAlphaFold2 - the

🚨 Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 🤯 It’s an open-source pipel…

Model ReleasesDGX agent

🚨 Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 🤯 It’s an open-source pipeline that replicates the exact daily loop of an ML researcher.

Construction of Knowledge Graph based on Language Model

ResearchDGX agent

arXiv:2604.19137v1 Announce Type: new Abstract: Knowledge Graph (KG) can effectively integrate valuable information from massive data, and thus has been rapidly developed and widely used in many field

← Previous
1…4849505152…203
Next →