AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlog
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,767 results
Model Releases

From Facts to Insights: A Persona-Driven Dual Memory Framework and Dataset for Role-Playing Agents

DGX agent

arXiv:2605.25693v1 Announce Type: new Abstract: While role-playing agents excel in short-term interactions, long-term conversations overwhelm context windows, motivating external memory frameworks. Cu

model-releasesarxiv-cs-cl
26 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Future-KL Regularized GRPO: Process-Level Credit Assignment from f-Divergence Regularization

DGX agent

arXiv:2601.10201v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) is widely used for critic-free Large Language Model (LLM) post-training, but its KL regularization i

local-aiarxiv-cs-ai
26 May 2026
Research

Geometric Evolution Maps: Extracting Stable Concept Probes from Transformer Residual Streams

DGX agent

arXiv:2605.25848v1 Announce Type: cross Abstract: Concept probes extracted from transformer residual streams are only as reliable as the layer from which they are extracted. The common practice of pro

researcharxiv-cs-ai
26 May 2026
Safety

GeoSVG-RL: Geometry-Aware Reinforcement Learning for Layout-Constrained Text-to-SVG Diagram Generation

DGX agent

arXiv:2605.25447v1 Announce Type: new Abstract: Generating structured, editable diagrams remains a significant challenge for contemporary large language models, despite their proficiency in general-pu

safetyarxiv-cs-cl
26 May 2026
Model Releases

GroupTravelBench: Benchmarking LLM Agents on Multi-Person Travel Planning

DGX agent

arXiv:2605.25200v1 Announce Type: new Abstract: Travel planning is a realistic task for evaluating the planning and tool-use abilities of LLM agents. However, existing benchmarks typically assume only

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Guided Flow Matching for Forward and Inverse PDE Problems with Sparse Observations: Algorithm and Theory

DGX agent

arXiv:2605.25509v1 Announce Type: cross Abstract: Reconstructing PDE solutions from sparse observations is a core challenge in scientific computing. We present FM4PDE, a flow-matching generative frame

model-releasesarxiv-cs-lg
26 May 2026
Local Ai

Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents

DGX agent

arXiv:2605.24598v1 Announce Type: new Abstract: Large language model (LLM) agents excel at solving complex long-horizon tasks through autonomous interaction with environments. However, their real-worl

local-aiarxiv-cs-ai
26 May 2026
Model Releases

HiGraph: A Large-Scale Hierarchical Graph Dataset for Malware Analysis

DGX agent

arXiv:2509.02113v2 Announce Type: replace-cross Abstract: The advancement of graph-based malware analysis is critically limited by the absence of large-scale datasets that capture the inherent hierarc

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

DGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

I found this Wired article on AI fact-checking frustrating. It could have been about why we continue to need human fact checkers (talk to pe…

DGX agent

I found this Wired article on AI fact-checking frustrating. It could have been about why we continue to need human fact checkers (talk to people, use judgement, resolve conflict). Instead it is full o

model-releasesethan-mollick--x
26 May 2026
Model Releases

IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

DGX agent

arXiv:2605.24659v1 Announce Type: new Abstract: LLM-based agents are increasingly deployed for complex tasks requiring planning, tool use, and interaction with external services. Their reliance on unt

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Keep the Proof State Live: Snapshotting for Efficient Tactic Search in Lean 4

DGX agent

arXiv:2605.25556v1 Announce Type: cross Abstract: Automated theorem proving systems built on Lean 4 increasingly rely on parallel tactic search over partially specified proofs, such as those generated

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Kolmogorov-Arnold Fourier Networks

DGX agent

arXiv:2502.06018v3 Announce Type: replace-cross Abstract: Although Kolmogorov-Arnold-based interpretable networks (KANs) possess strong theoretical expressiveness, they suffer from severe parameter ex

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

LLMs on prem. but the prem is your wardrobe. http://shop.cohere.com The design team cooked on this one.

DGX agent

Cohere has announced a creative marketing initiative featuring LLM-themed merchandise available through their shop, with the tagline playing on the phrase 'on premises' to humorously suggest their lan

model-releasescohere--x
26 May 2026
Local Ai

Lngram: N-gram Conditional Memory in Latent Space

DGX agent

arXiv:2605.24869v1 Announce Type: new Abstract: Sequence modeling requires both compositional reasoning and local static knowledge retrieval, yet standard Transformers handle both through dense comput

local-aiarxiv-cs-cl
26 May 2026
Applications

LWM-CDE: A Representation Space for Wireless Data Reasoning and Transferability

DGX agent

arXiv:2605.24077v1 Announce Type: cross Abstract: Machine learning deployments in real-world wireless communication tasks face significant generalization challenges due to location and environment-spe

applicationsarxiv-cs-lg
26 May 2026
Model Releases

Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework

DGX agent

arXiv:2605.24661v1 Announce Type: new Abstract: LLMs have achieved remarkable success in complex reasoning tasks, yet current evaluation approaches predominantly rely on final-answer correctness, offe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Memory-Induced Tool-Drift in LLM Agents

DGX agent

arXiv:2605.24941v1 Announce Type: cross Abstract: Modern LLM agents combine long-term memory for personalization with tool-calling interfaces for taking actions in the world -- a combination underpinn

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

MIND: Multi-Scale Intent Diffusion for Text-Driven Physics-Based Humanoid Control

DGX agent

arXiv:2605.26006v1 Announce Type: cross Abstract: Enabling physics-based humanoids to execute diverse behaviors from high-level textual commands remains a significant challenge. Existing methods typic

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

MindAlign: Bridging EEG, Vision, and Language for Zero-Shot Visual Decoding

DGX agent

arXiv:2605.24523v1 Announce Type: cross Abstract: Visual decoding from brain signals is a key challenge at the intersection of computer vision and neuroscience, requiring methods that bridge neural re

model-releasesarxiv-cs-cl
26 May 2026
Tutorials

Multitask learning with semiempirical orbital charges enables sample-efficient MLIPs

DGX agent

arXiv:2605.24073v1 Announce Type: cross Abstract: Machine learning interatomic potentials (MLIPs) require generating computationally expensive, large-scale training datasets to accurately simulate mat

tutorialsarxiv-cs-lg
26 May 2026
Model Releases

Muon in Associative Memory Learning: Training Dynamics and Scaling Laws

DGX agent

arXiv:2602.05725v2 Announce Type: replace Abstract: Muon updates matrix parameters via the matrix sign of the gradient and has shown strong empirical gains, yet its dynamics and scaling behavior remai

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Nine reasons why I warned OpenAI might fail and turn out to be the WeWork of AI, from two years ago.

DGX agent

Nine reasons why I warned OpenAI might fail and turn out to be the WeWork of AI, from two years ago. 9 reasons that OpenAI could someday be seen as the WeWork of AI: 👉 Lots of competitors are catching

model-releasesgary-marcus--x
26 May 2026
Model Releases

Noise-Robust Financial Numerical Entity Attribute Tagging

DGX agent

arXiv:2605.24910v1 Announce Type: new Abstract: Financial Numerical Entity (FNE) understanding aims to recover the meaning of numerical mentions in financial reports. Existing studies primarily focus

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

NormimesDirection: Restoring the Missing Query Norm in Vision Linear Attention

DGX agent

arXiv:2506.21137v3 Announce Type: replace Abstract: Linear attention mitigates the quadratic complexity of softmax attention but suffers from a critical loss of expressiveness. We identify two primary

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

On the Epistemic Uncertainty of Overparametrized Neural Networks

DGX agent

arXiv:2605.25234v1 Announce Type: cross Abstract: Epistemic uncertainty is often viewed as a reducible uncertainty that vanishes with increasing data. This perspective implicitly assumes parameter ide

model-releasesarxiv-cs-ai
26 May 2026
Hardware

Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers

DGX agent

arXiv:2605.25346v1 Announce Type: cross Abstract: Neural network (NN) dynamics models and control policies achieve strong performance in robotics, but providing sound guarantees under uncertainty rema

hardwarearxiv-cs-ai
26 May 2026
Model Releases

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

DGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

model-releasesarxiv-cs-ai
26 May 2026
Research

{Phi}-Noise: Training-Free Temporal Video Conditioning via Phase-Based Noise Manipulation

DGX agent

arXiv:2605.24509v1 Announce Type: cross Abstract: Latent video diffusion models generate videos by progressively transforming Gaussian noise into realistic samples conditioned on text or visual inputs

researcharxiv-cs-ai
26 May 2026
Research

Poisoning the Watchtower: Prompt Injection Attacks Against LLM-Augmented Security Operations Through Adversarial Log Content

DGX agent

arXiv:2605.24421v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as analyst assistants in security operations centers (SOCs), where they ingest log and alert data t

researcharxiv-cs-lg
26 May 2026
Model Releases

Position: AI for Science Should Treat Measurement-to-Dataset Pipelines as Inference Components

DGX agent

arXiv:2605.24558v1 Announce Type: new Abstract: AI for Science (AI4Science) workflows often treat the released dataset as a fixed interface to the underlying system. However, in domains relying on ind

model-releasesarxiv-cs-lg
26 May 2026
Research

PowLU: An Activation Function for Stable Pre-Training of LLMs

DGX agent

arXiv:2605.25704v1 Announce Type: new Abstract: In contemporary large language models (LLMs), the swish-gated linear unit (SwiGLU) activation function is widely adopted to regulate the information flo

researcharxiv-cs-cl
26 May 2026
Local Ai

Prefix Teach, Suffix Fade: Local Teachability Collapse in Strong-to-Weak On-Policy Distillation

DGX agent

arXiv:2605.13643v2 Announce Type: replace Abstract: On-policy distillation (OPD) trains a student model on its own rollouts using dense feedback from a stronger teacher. Prior literature suggests that

local-aiarxiv-cs-cl
26 May 2026
Model Releases

Querying structural and functional niches on spatial transcriptomics data

DGX agent

arXiv:2410.10652v4 Announce Type: replace-cross Abstract: Cells in multicellular organisms coordinate to form structural and functional niches. With spatial transcriptomics (ST) enabling gene expressi

model-releasesarxiv-cs-lg
26 May 2026
Safety

Reasoning as an Attack Surface: Adaptive Evolutionary CoT Jailbreaks for LLMs

DGX agent

arXiv:2605.24497v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated remarkable capabilities in reasoning and generation tasks and are increasingly deployed in real-world ap

safetyarxiv-cs-ai
26 May 2026
Model Releases

RED: Adaptive Real-Time DAG Scheduling for Robotic Inference under Environmental Dynamics

DGX agent

arXiv:2605.24044v1 Announce Type: new Abstract: Robots deployed in dynamic environments must contend with environment-driven changes that reshape computation at runtime: new tasks may appear, preceden

model-releasesarxiv-cs-ro
26 May 2026
Safety

Reinforcement Learning from Denoising Feedback

DGX agent

arXiv:2605.25638v1 Announce Type: new Abstract: Policy loss estimation remains a fundamental and long-standing challenge in reinforcement learning (RL) for diffusion language models (dLLMs). We introd

safetyarxiv-cs-cl
26 May 2026
Research

Remote sensing data imputation using deep learning for multispectral imagery

DGX agent

arXiv:2605.24003v1 Announce Type: cross Abstract: Remote sensing techniques have been increasingly utilised in aquatic applications in recent years. A common challenge in using optical satellite data

researcharxiv-cs-ai
26 May 2026
Safety

Rewarding Structural Conformance of Reasoning using Process Mining

DGX agent

arXiv:2510.25065v3 Announce Type: replace Abstract: Recent advances in sparse reward policy gradient methods have enabled effective reinforcement learning (RL)-based language model post-training. Howe

safetyarxiv-cs-ai
26 May 2026
Model Releases

Safety Generalization Under Distribution Shift in Safe Reinforcement Learning: A Diabetes Testbed

DGX agent

arXiv:2601.21094v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning (RL) algorithms are typically evaluated under fixed training conditions. We investigate whether training-time safe

model-releasesarxiv-cs-ai
26 May 2026
Agents

SAM: State-Adaptive Memory for Long-Horizon Reasoning Agent

DGX agent

arXiv:2605.24468v1 Announce Type: new Abstract: Long-horizon agentic reasoning requires large language models to act over long interaction histories containing thoughts, tool calls, observations, and

agentsarxiv-cs-ai
26 May 2026
Research

SentGraph: Hierarchical Sentence Graph for Multi-hop Retrieval-Augmented Question Answering

DGX agent

arXiv:2601.03014v3 Announce Type: replace-cross Abstract: Traditional Retrieval-Augmented Generation (RAG) effectively supports single-hop question answering with large language models but faces signi

researcharxiv-cs-ai
26 May 2026
Model Releases

'Si'multaneous 'S'patial-'T'emporal Message Passing for Dynamic Graph Representation Learning

DGX agent

arXiv:2605.25548v1 Announce Type: cross Abstract: Dynamic graph neural networks (DGNNs) that operate on snapshot sequences typically fall into one of two categories. Temporal-first approaches build pe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Smart Timing for Mining: A Deep Learning Framework for Bitcoin Hardware ROI Prediction

DGX agent

arXiv:2512.05402v2 Announce Type: replace-cross Abstract: Bitcoin mining hardware acquisition requires strategic timing due to volatile markets, rapid technological obsolescence, and protocol-driven r

model-releasesarxiv-cs-ai
26 May 2026
Agents

SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharing

DGX agent

arXiv:2512.24008v3 Announce Type: replace Abstract: Personalized search demands the ability to model users' evolving, multi-dimensional information needs; a challenge for systems constrained by static

agentsarxiv-cs-ai
26 May 2026
Research

Spiking the training data to correct for test set contamination

DGX agent

arXiv:2605.24818v1 Announce Type: cross Abstract: The literature on test set contamination largely focuses on detection, but the correction of contaminated test scores is underexplored. Our core propo

researcharxiv-cs-cl
26 May 2026
Safety

StakeBench: Evaluating Language Understanding Grounded in Market Commitment

DGX agent

arXiv:2605.26074v1 Announce Type: cross Abstract: Existing financial NLP benchmarks often rely on labels supplied by outside observers, measuring how language is perceived rather than what speakers ha

safetyarxiv-cs-ai
26 May 2026
Applications

Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions

DGX agent

arXiv:2605.24452v1 Announce Type: cross Abstract: Legal NLP benchmarks evaluate models on randomly split data, implicitly assuming that legal language is stationary. We test this assumption by fine-tu

applicationsarxiv-cs-ai
26 May 2026
← Previous
1…712713714715716…1371
Next →