AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

Efficient Benchmarking Is Just Feature Selection and Multiple Regression

DGX agent

arXiv:2605.25773v1 Announce Type: cross Abstract: Efficient benchmarking techniques aim to lower the computational cost of evaluating LLMs by predicting full benchmark scores using only a subset of a

model-releasesarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EMA-Nesterov: Stabilizing Nesterov's Lookahead for Accelerated Deep Learning Optimization

DGX agent

arXiv:2605.25395v1 Announce Type: new Abstract: Lookahead-based acceleration methods, such as Nesterov's momentum, are widely used in optimization, but they often become unreliable in deep learning tr

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Emission-Aware Reinforcement Learning for Sustainable Electric Vehicle Charging and Carbon Dioxide Reduction Under Varying Renewable Penetration

DGX agent

arXiv:2605.24543v1 Announce Type: new Abstract: The rapid growth of Electric Vehicle (EV) adoption challenges power distribution networks through peak load spikes, voltage instability, and transformer

model-releasesarxiv-cs-ai
26 May 2026
Research

End-to-End Intracortical Speech Decoding from Neural Activity

DGX agent

arXiv:2605.24313v1 Announce Type: new Abstract: Current high-performing intracortical speech neuroprostheses achieve low word error rates but typically rely on external language models during inferenc

researcharxiv-cs-cl
26 May 2026
Applications

Equip Pre-ranking with Target Attention by Residual Quantization

DGX agent

arXiv:2509.16931v3 Announce Type: replace-cross Abstract: The pre-ranking stage in industrial recommendation systems faces a fundamental conflict between efficiency and effectiveness. While powerful m

applicationsarxiv-cs-ai
26 May 2026
Research

EVA: Accelerating LLM Decoding via an Efficient Vector Quantization Architecture

DGX agent

arXiv:2605.24144v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved impressive performance across diverse domains but remain inefficient during the autoregressive decoding pha

researcharxiv-cs-lg
26 May 2026
Model Releases

everybody talks about the china->us catchup not enough people talking about the us-> china catchup great job @o_lacombe et al, @robert_mchar…

DGX agent

everybody talks about the china->us catchup not enough people talking about the us-> china catchup great job @o_lacombe et al, @robert_mchardy et al! [AINews 3 Apr 2026] Gemma 4: The world's best smal

model-releasesswyx--x
26 May 2026
Model Releases

EvoEGF-Mol: Evolving Exponential Geodesic Flow for Structure-based Drug Design

DGX agent

arXiv:2601.22466v2 Announce Type: replace Abstract: Structure-Based Drug Design (SBDD) aims to discover bioactive ligands. Conventional approaches construct probability paths separately in Euclidean a

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Explore Before You Solve: The Speed--Depth Trade-off in Epistemic Agents for ARC-AGI-3

DGX agent

arXiv:2605.25931v1 Announce Type: new Abstract: We systematically investigate all 25 public ARC-AGI-3 games and find that every one is reachable through non-intelligent strategies: 10 in a single blin

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Extending Embodied Question Answering from Perception to Decision

DGX agent

arXiv:2605.25813v1 Announce Type: new Abstract: Embodied Question Answering (EQA) connects perception, reasoning, and interaction within embodied environments. However, existing datasets and benchmark

model-releasesarxiv-cs-ro
26 May 2026
Tutorials

Factorize to Generalize: Retrieval-Guided Invariant-Dynamic Decomposition for Time Series Forecasting

DGX agent

arXiv:2605.24911v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have recently achieved strong zero-shot forecasting performance through large-scale pretraining and retrieval-au

tutorialsarxiv-cs-ai
26 May 2026
Safety

Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning

DGX agent

arXiv:2605.24286v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning is useful for monitoring language models only when the reasoning trace faithfully reflects the computation that produ

safetyarxiv-cs-cl
26 May 2026
Local Ai

False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs

DGX agent

arXiv:2510.14925v4 Announce Type: replace Abstract: High-confidence errors in large language models are often treated as fragile failures. We study an alternative: some errors may be false fixed point

local-aiarxiv-cs-ai
26 May 2026
Research

Feature Resemblance: Towards a Theoretical Understanding of Analogical Reasoning in Transformers

DGX agent

arXiv:2603.05143v3 Announce Type: replace Abstract: Understanding reasoning in large language models is complicated by evaluations that conflate multiple reasoning types. We isolate analogical reasoni

researcharxiv-cs-cl
26 May 2026
Model Releases

FLOATBench: A Dataset and Benchmark for Floating Offshore Wind Turbine Tower Fatigue

DGX agent

arXiv:2605.25717v1 Announce Type: new Abstract: Most of the world's offshore wind resource lies in waters too deep for fixed-bottom foundations, making floating offshore wind turbines (FOWTs) essentia

model-releasesarxiv-cs-ai
26 May 2026
Safety

From Automation to Collaboration: Human-in-the-Loop Methods for Safe and Trustworthy NLP

DGX agent

arXiv:2605.25226v1 Announce Type: new Abstract: Large language models are widely deployed in high-stakes NLP tasks, yet risks such as bias, hallucination, adversarial vulnerability and unreliable gene

safetyarxiv-cs-cl
26 May 2026
Model Releases

From Facts to Insights: A Persona-Driven Dual Memory Framework and Dataset for Role-Playing Agents

DGX agent

arXiv:2605.25693v1 Announce Type: new Abstract: While role-playing agents excel in short-term interactions, long-term conversations overwhelm context windows, motivating external memory frameworks. Cu

model-releasesarxiv-cs-cl
26 May 2026
Local Ai

Future-KL Regularized GRPO: Process-Level Credit Assignment from f-Divergence Regularization

DGX agent

arXiv:2601.10201v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) is widely used for critic-free Large Language Model (LLM) post-training, but its KL regularization i

local-aiarxiv-cs-ai
26 May 2026
Research

Geometric Evolution Maps: Extracting Stable Concept Probes from Transformer Residual Streams

DGX agent

arXiv:2605.25848v1 Announce Type: cross Abstract: Concept probes extracted from transformer residual streams are only as reliable as the layer from which they are extracted. The common practice of pro

researcharxiv-cs-ai
26 May 2026
Safety

GeoSVG-RL: Geometry-Aware Reinforcement Learning for Layout-Constrained Text-to-SVG Diagram Generation

DGX agent

arXiv:2605.25447v1 Announce Type: new Abstract: Generating structured, editable diagrams remains a significant challenge for contemporary large language models, despite their proficiency in general-pu

safetyarxiv-cs-cl
26 May 2026
Model Releases

GroupTravelBench: Benchmarking LLM Agents on Multi-Person Travel Planning

DGX agent

arXiv:2605.25200v1 Announce Type: new Abstract: Travel planning is a realistic task for evaluating the planning and tool-use abilities of LLM agents. However, existing benchmarks typically assume only

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory

DGX agent

arXiv:2605.25509v1 Announce Type: cross Abstract: Reconstructing PDE solutions from sparse observations is a core challenge in scientific computing. We present FM4PDE, a flow-matching generative frame

model-releasesarxiv-cs-lg
26 May 2026
Local Ai

Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents

DGX agent

arXiv:2605.24598v1 Announce Type: new Abstract: Large language model (LLM) agents excel at solving complex long-horizon tasks through autonomous interaction with environments. However, their real-worl

local-aiarxiv-cs-ai
26 May 2026
Model Releases

HiGraph: A Large-Scale Hierarchical Graph Dataset for Malware Analysis

DGX agent

arXiv:2509.02113v2 Announce Type: replace-cross Abstract: The advancement of graph-based malware analysis is critically limited by the absence of large-scale datasets that capture the inherent hierarc

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

DGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

I found this Wired article on AI fact-checking frustrating. It could have been about why we continue to need human fact checkers (talk to pe…

DGX agent

I found this Wired article on AI fact-checking frustrating. It could have been about why we continue to need human fact checkers (talk to people, use judgement, resolve conflict). Instead it is full o

model-releasesethan-mollick--x
26 May 2026
Model Releases

IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

DGX agent

arXiv:2605.24659v1 Announce Type: new Abstract: LLM-based agents are increasingly deployed for complex tasks requiring planning, tool use, and interaction with external services. Their reliance on unt

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Keep the Proof State Live: Snapshotting for Efficient Tactic Search in Lean 4

DGX agent

arXiv:2605.25556v1 Announce Type: cross Abstract: Automated theorem proving systems built on Lean 4 increasingly rely on parallel tactic search over partially specified proofs, such as those generated

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Kolmogorov-Arnold Fourier Networks

DGX agent

arXiv:2502.06018v3 Announce Type: replace-cross Abstract: Although Kolmogorov-Arnold-based interpretable networks (KANs) possess strong theoretical expressiveness, they suffer from severe parameter ex

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

LLMs on prem. but the prem is your wardrobe. http://shop.cohere.com The design team cooked on this one.

DGX agent

Cohere has announced a creative marketing initiative featuring LLM-themed merchandise available through their shop, with the tagline playing on the phrase 'on premises' to humorously suggest their lan

model-releasescohere--x
26 May 2026
Local Ai

Lngram: N-gram Conditional Memory in Latent Space

DGX agent

arXiv:2605.24869v1 Announce Type: new Abstract: Sequence modeling requires both compositional reasoning and local static knowledge retrieval, yet standard Transformers handle both through dense comput

local-aiarxiv-cs-cl
26 May 2026
Applications

LWM-CDE: A Representation Space for Wireless Data Reasoning and Transferability

DGX agent

arXiv:2605.24077v1 Announce Type: cross Abstract: Machine learning deployments in real-world wireless communication tasks face significant generalization challenges due to location and environment-spe

applicationsarxiv-cs-lg
26 May 2026
Model Releases

Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework

DGX agent

arXiv:2605.24661v1 Announce Type: new Abstract: LLMs have achieved remarkable success in complex reasoning tasks, yet current evaluation approaches predominantly rely on final-answer correctness, offe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Memory-Induced Tool-Drift in LLM Agents

DGX agent

arXiv:2605.24941v1 Announce Type: cross Abstract: Modern LLM agents combine long-term memory for personalization with tool-calling interfaces for taking actions in the world -- a combination underpinn

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

MIND: Multi-Scale Intent Diffusion for Text-Driven Physics-Based Humanoid Control

DGX agent

arXiv:2605.26006v1 Announce Type: cross Abstract: Enabling physics-based humanoids to execute diverse behaviors from high-level textual commands remains a significant challenge. Existing methods typic

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

MindAlign: Bridging EEG, Vision, and Language for Zero-Shot Visual Decoding

DGX agent

arXiv:2605.24523v1 Announce Type: cross Abstract: Visual decoding from brain signals is a key challenge at the intersection of computer vision and neuroscience, requiring methods that bridge neural re

model-releasesarxiv-cs-cl
26 May 2026
Tutorials

Multitask learning with semiempirical orbital charges enables sample-efficient MLIPs

DGX agent

arXiv:2605.24073v1 Announce Type: cross Abstract: Machine learning interatomic potentials (MLIPs) require generating computationally expensive, large-scale training datasets to accurately simulate mat

tutorialsarxiv-cs-lg
26 May 2026
Model Releases

Muon in Associative Memory Learning: Training Dynamics and Scaling Laws

DGX agent

arXiv:2602.05725v2 Announce Type: replace Abstract: Muon updates matrix parameters via the matrix sign of the gradient and has shown strong empirical gains, yet its dynamics and scaling behavior remai

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Nine reasons why I warned OpenAI might fail and turn out to be the WeWork of AI, from two years ago.

DGX agent

Nine reasons why I warned OpenAI might fail and turn out to be the WeWork of AI, from two years ago. 9 reasons that OpenAI could someday be seen as the WeWork of AI: 👉 Lots of competitors are catching

model-releasesgary-marcus--x
26 May 2026
Model Releases

Noise-Robust Financial Numerical Entity Attribute Tagging

DGX agent

arXiv:2605.24910v1 Announce Type: new Abstract: Financial Numerical Entity (FNE) understanding aims to recover the meaning of numerical mentions in financial reports. Existing studies primarily focus

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

NormimesDirection: Restoring the Missing Query Norm in Vision Linear Attention

DGX agent

arXiv:2506.21137v3 Announce Type: replace Abstract: Linear attention mitigates the quadratic complexity of softmax attention but suffers from a critical loss of expressiveness. We identify two primary

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

On the Epistemic Uncertainty of Overparametrized Neural Networks

DGX agent

arXiv:2605.25234v1 Announce Type: cross Abstract: Epistemic uncertainty is often viewed as a reducible uncertainty that vanishes with increasing data. This perspective implicitly assumes parameter ide

model-releasesarxiv-cs-ai
26 May 2026
Hardware

Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers

DGX agent

arXiv:2605.25346v1 Announce Type: cross Abstract: Neural network (NN) dynamics models and control policies achieve strong performance in robotics, but providing sound guarantees under uncertainty rema

hardwarearxiv-cs-ai
26 May 2026
Model Releases

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

DGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

model-releasesarxiv-cs-ai
26 May 2026
Research

{Phi}-Noise: Training-Free Temporal Video Conditioning via Phase-Based Noise Manipulation

DGX agent

arXiv:2605.24509v1 Announce Type: cross Abstract: Latent video diffusion models generate videos by progressively transforming Gaussian noise into realistic samples conditioned on text or visual inputs

researcharxiv-cs-ai
26 May 2026
Research

Poisoning the Watchtower: Prompt Injection Attacks Against LLM-Augmented Security Operations Through Adversarial Log Content

DGX agent

arXiv:2605.24421v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as analyst assistants in security operations centers (SOCs), where they ingest log and alert data t

researcharxiv-cs-lg
26 May 2026
Model Releases

Position: AI for Science Should Treat Measurement-to-Dataset Pipelines as Inference Components

DGX agent

arXiv:2605.24558v1 Announce Type: new Abstract: AI for Science (AI4Science) workflows often treat the released dataset as a fixed interface to the underlying system. However, in domains relying on ind

model-releasesarxiv-cs-lg
26 May 2026
Research

PowLU: An Activation Function for Stable Pre-Training of LLMs

DGX agent

arXiv:2605.25704v1 Announce Type: new Abstract: In contemporary large language models (LLMs), the swish-gated linear unit (SwiGLU) activation function is widely adopted to regulate the information flo

researcharxiv-cs-cl
26 May 2026
← Previous
1…712713714715716…1371
Next →