AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
26 May 2026

Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models

SafetyDGX agent

arXiv:2605.24564v1 Announce Type: new Abstract: Backtesting large language models (LLMs) on historical financial data is unreliable because pre-training cuts off after the events happened. An LLM trai

SURGE: On the Potential of Large Language Models as General-Purpose Surrogate Code Executors

Model ReleasesDGX agent

arXiv:2502.11167v5 Announce Type: replace-cross Abstract: Neural surrogate models are powerful and efficient tools in data mining. Meanwhile, large language models (LLMs) have demonstrated remarkable

Synheart Capacity: A Theory-Driven Physiological Representation of Cognitive Capacity Dynamics from Wearable Signals

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.24416v1 Announce Type: new Abstract: Human cognitive performance is constrained by limited mental resources, yet continuous computational estimation of cognitive capacity dynamics remains a

T2S-MPC: Time-Embedded Online Adaptive Model Predictive Control for Time-Varying Dynamics

ResearchDGX agent

arXiv:2605.24852v1 Announce Type: new Abstract: Recent advances in learning-based model predictive control (MPC) have leveraged neural networks for online model learning, achieving strong performance

TaBIIC2: Interactive Building of Ontological Taxonomies using Weighted Self-Organizing Maps

ResearchDGX agent

arXiv:2605.24899v1 Announce Type: new Abstract: Ontologies represent the conceptual knowledge of a domain. At the core of an ontology is the taxonomy of concepts and subconcepts that represent specifi

TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

SafetyDGX agent

arXiv:2605.25547v1 Announce Type: new Abstract: Existing embodied control research demonstrates remarkable performance improvements by scaling training data and model size. We instead explore inferenc

Task-Aligned Self-Supervised Learning for Medical Image Analysis: A Systematic Review and Practical Design Guidelines

SafetyDGX agent

arXiv:2605.23995v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has emerged as a promising paradigm for addressing the annotation bottleneck in medical imaging by learning representat

Teaching large language models to reason like expert diagnosticians

Model ReleasesDGX agent

arXiv:2509.12194v2 Announce Type: replace Abstract: Differential diagnosis is an iterative process that integrates patient information with broader medical knowledge. Clinical case series such as the

Teaching Through Analogies: A Modular Pipeline for Educational Analogy Generation

Model ReleasesDGX agent

arXiv:2605.24211v1 Announce Type: cross Abstract: Analogies help learners understand unfamiliar concepts by relating them to known concepts. Despite recent advances, large language models (LLMs) conti

Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions

ApplicationsDGX agent

arXiv:2605.24452v1 Announce Type: cross Abstract: Legal NLP benchmarks evaluate models on randomly split data, implicitly assuming that legal language is stationary. We test this assumption by fine-tu

Temporal Score Rescaling for Temperature Sampling in Diffusion and Flow Models

Local AiDGX agent

arXiv:2510.01184v2 Announce Type: replace Abstract: We present a mechanism to steer the sampling diversity of denoising diffusion and flow matching models, allowing users to sample from a sharper or b

Terrain-Adaptive Grouser Wheel for Optimal Planetary Exploration: Design and Experimental Investigation

ResearchDGX agent

arXiv:2605.24311v1 Announce Type: new Abstract: Planetary rovers operating in extraterrestrial environments often encounter significant mobility challenges due to varying terrain features such as grad

Test-Time Deep Thinking to Explore Implicit Rules

AgentsDGX agent

arXiv:2605.24828v1 Announce Type: new Abstract: With the continuous advancement of Large Language Models (LLMs), intelligent agents are becoming increasingly vital. However, these agents often fail in

Test-Time Graph Search for Goal-Conditioned Reinforcement Learning

Model ReleasesDGX agent

arXiv:2510.07257v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning (GCRL) often struggles with long-horizon tasks, where errors in value estimation accumulate and prod

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation

Model ReleasesDGX agent

arXiv:2605.25488v1 Announce Type: cross Abstract: Audio-driven talking-head generation has achieved remarkable progress with recent models such as AniTalker, FLOAT, and Sonic. Despite their success, m

Testing the Deliteralization Hypothesis in Human and Machine Translation

ResearchDGX agent

arXiv:2605.25686v1 Announce Type: new Abstract: The recent shift from dedicated NMT systems to general-purpose LLMs has reshaped machine translation, with LLMs reported to produce more fluent, less li

TGFormer: Towards Temporal Graph Transformer with Auto-Correlation Mechanism

ResearchDGX agent

arXiv:2605.24971v1 Announce Type: cross Abstract: The growing interest in Temporal Graph Neural Networks (TGNNs) stems from their ability to model complex dynamics and deliver superior performance. Ho

Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization

ResearchDGX agent

arXiv:2605.25928v1 Announce Type: new Abstract: We describe the winning system for Task 2 of the KSAA-2026 Shared Task on Arabic Speech Dictation with Automatic Diacritization. The task requires produ

The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models

Model ReleasesDGX agent

arXiv:2605.25510v1 Announce Type: new Abstract: Children increasingly have access to Large Language Models (LLMs), which may expose them to responses that are developmentally inappropriate or require

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible

SafetyDGX agent

arXiv:2605.25739v1 Announce Type: new Abstract: We prove that no reinforcement learning policy with confidence-gated autonomy can simultaneously achieve maximum helpfulness, optimal calibration, and f

The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth

SafetyDGX agent

arXiv:2605.24856v1 Announce Type: cross Abstract: Concept formation in transformer language models is depth-extended, not a single-layer event: concepts emerge gradually across a contiguous region of

The Impact of Large Language Models on Open-source Innovation: Evidence from GitHub Copilot

ResearchDGX agent

arXiv:2409.08379v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are reshaping knowledge work, yet their impact on voluntary, self-guided open innovation forums (contributors cho

The Implicit Bias of Adam and Muon on Smooth Homogeneous Neural Networks

SafetyDGX agent

arXiv:2602.16340v3 Announce Type: replace Abstract: We study the implicit bias of momentum-based optimizers on smooth homogeneous models. We show that extit{momentum steepest descent} algorithms like

The LSCD Benchmark: a Testbed for Diachronic Word Meaning Tasks

Model ReleasesDGX agent

arXiv:2404.00176v3 Announce Type: replace Abstract: Lexical Semantic Change Detection (LSCD) is a complex, lemma-level task, which is usually operationalized based on two subsequently applied usage-le

The Many Faces of On-Policy Distillation: Pitfalls, Mechanisms, and Fixes

SafetyDGX agent

arXiv:2605.11182v2 Announce Type: replace Abstract: On-policy distillation (OPD) and on-policy self-distillation (OPSD) have emerged as promising post-training methods for large language models, offer

The meaning of prompts and the prompts of meaning: Semiotic reflections and modelling

ResearchDGX agent

arXiv:2509.14250v2 Announce Type: replace Abstract: This paper explores prompts and prompting in large language models (LLMs) as dynamic semiotic phenomena, drawing on Peirce's triadic model of signs,

The Meme Is the Message: Generative Memesis and AI Visuals in the 2024 USA Presidential Elections

ApplicationsDGX agent

arXiv:2411.00934v2 Announce Type: replace-cross Abstract: Visual content on social media has become increasingly influential in shaping political discourse and civic engagement, but it also limits par

The Model Is Not the Product: A Dual-Pillar Architecture for Local-First Psychological Coaching

Model ReleasesDGX agent

arXiv:2605.24411v1 Announce Type: new Abstract: Existing language model applications struggle to meet the demand for emotionally oriented support, primarily due to their inability to maintain deep, pe

The Model Parking Tax: Quantifying the Hidden Energy Cost of Always-On GPU Model Deployment

Local AiDGX agent

arXiv:2605.23918v1 Announce Type: cross Abstract: The AI inference industry keeps models loaded in GPU memory around the clock to avoid cold-start latency, implicitly treating idle power as a fixed co

The Multilingual Curse at the Retrieval Layer: Evidence from Amharic

ResearchDGX agent

arXiv:2605.24556v1 Announce Type: cross Abstract: Multilingual retrieval increasingly underpins cross-lingual question answering and retrieval-augmented generation. Strong zero-shot scores on multilin

The Normalized Maximum Likelihood for Regular Non-Smooth Models: Measure-Theoretic Foundations and Geometric Sampling

ResearchDGX agent

arXiv:2605.24477v1 Announce Type: new Abstract: The Normalized Maximum Likelihood (NML) codelength, or stochastic complexity, represents a principled criterion for universal coding. While recent coare

The Path Matters: Learning a Token-Commitment Policy for Diffusion Language Models

SafetyDGX agent

arXiv:2605.24697v1 Announce Type: cross Abstract: Diffusion large language models promise faster generation by refining many token positions in parallel, but this parallelism introduces a hidden contr

The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench

Model ReleasesDGX agent

arXiv:2605.24782v1 Announce Type: new Abstract: While Vision Foundation Models (VFMs) excel at predictive tasks on satellite imagery, their performance can arise from visual correlations rather than u

The Quantization Benefits of Residual-Free Transformers

ResearchDGX agent

arXiv:2605.25880v1 Announce Type: new Abstract: Large-scale transformer training and deployment are increasingly constrained by the transfer of activations, gradients, and optimizer states across acce

The Time is Here for Just-in-Time Systems: Challenges and Opportunities

Model ReleasesDGX agent

arXiv:2605.24096v1 Announce Type: cross Abstract: Core systems like key-value stores have historically taken years to build, and are designed to be general so as to amortize cost across deployments, p

The Timing Dependencies of Trust: Speed, Accuracy, and cBCI Neuro-Decoupling in Human-AI Teams

ResearchDGX agent

arXiv:2605.25868v1 Announce Type: cross Abstract: The speed and accuracy of an artificial teammate fundamentally alter the failure states of Human-AI integration. While high-speed AI interventions ris

The Tokenizer Tax Across 25 European Languages: Domain Invariance, Cross-Lingual Few-Shot Effects, and the Ukrainian Penalty

ResearchDGX agent

arXiv:2605.24718v1 Announce Type: new Abstract: Tokenizer fertility the number of tokens per word imposes a hidden cost on non-English NLP. We measure fertility for ten foundation models across 25 Eur

Theoretical Analysis of Sparse Optimization with Reparameterization, Weight Decay, and Adaptive Learning Rate

ResearchDGX agent

arXiv:2605.25134v1 Announce Type: cross Abstract: Sparse optimization is a fundamental challenge in various practical applications. A popular approach to sparse optimization is ell_p regularization. H

They Are Not the Same: Direct Causes Are Not Grounded Emotion Explanations

ResearchDGX agent

arXiv:2605.25208v1 Announce Type: new Abstract: Emotion-Cause Pair Extraction (ECPE) was introduced to explain why an emotion occurs, but this goal is now often reduced to binary pair/non-pair predict

TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning

Model ReleasesDGX agent

arXiv:2605.25850v1 Announce Type: cross Abstract: This paper investigates large language model (LLM) abstention learning, specifically using ternary reward, which incentivize truthfulness in large lan

TIGER: Text-Informed Generalized Enzyme-Reaction Retrieval

ResearchDGX agent

arXiv:2605.24489v1 Announce Type: new Abstract: Enzyme-reaction retrieval is a fundamental problem in computational biology, underpinning enzyme characterization, reaction mechanism elucidation, and t

TimeSpot: Benchmarking Geo-Temporal Understanding in Vision-Language Models in Real-World Settings

Model ReleasesDGX agent

arXiv:2603.06687v2 Announce Type: replace-cross Abstract: Geo-temporal understanding, the ability to infer location, time, and contextual properties from visual input alone, underpins applications suc

Tiny Brains, Giant Impact: Uncovering the Keystone Neurons of LLM with Just a Few Prompts

Model ReleasesDGX agent

arXiv:2605.24846v1 Announce Type: cross Abstract: Large language models (LLMs) display strong comprehensive abilities, yet the internal mechanisms that support these behaviors remain insufficiently un

TinyFormer: Preserving Tiny Objects in YOLO-DETRHybridReal-time Detectors

ResearchDGX agent

arXiv:2605.25046v1 Announce Type: cross Abstract: YOLO-series and DETR-based detectors struggle with tiny-object detection. YOLO-style models benefit from efficient dense prediction, but their large-s

Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams

AgentsDGX agent

arXiv:2605.25310v1 Announce Type: new Abstract: Tool-using LLM agents produce trajectories whose calls form a directed dependency graph: earlier tool outputs supply arguments to later calls. Whether t

ToolRegistry: A Protocol-Agnostic Tool Management Library for Function-Calling LLMs

Model ReleasesDGX agent

arXiv:2507.10593v3 Announce Type: replace-cross Abstract: Every LLM tool call is structurally an RPC -- a function name, JSON arguments, and a serialized result -- yet each protocol (native Python, MC

TopoAlign: Topology-Aware Visual Representation Alignment

SafetyDGX agent

arXiv:2605.25541v1 Announce Type: cross Abstract: Neural networks encode inputs as high-dimensional vectors, known as representations, that capture how models process data by encoding task-relevant st

Topology-Driven Transferability Estimation of Medical Foundation Models for Segmentation

Model ReleasesDGX agent

arXiv:2602.23916v2 Announce Type: replace-cross Abstract: The advent of large-scale self-supervised learning (SSL) has produced a vast zoo of medical foundation models. However, selecting optimal medi

TorchLean: Formalizing Neural Networks in Lean

SafetyDGX agent

arXiv:2602.22631v2 Announce Type: replace-cross Abstract: Neural networks are increasingly deployed in scientific, safety critical, and mission critical pipelines, yet verification and analysis are of

Toward a Benchmark for Controllable Simulation of Imperfect Students with Large Language Models

Model ReleasesDGX agent

arXiv:2605.25601v1 Announce Type: cross Abstract: Teacher education requires deliberate practice with learners who exhibit identifiable strengths, weaknesses, and partial mastery. Large language model

Toward Enactive Artificial Intelligence

AgentsDGX agent

arXiv:2605.24238v1 Announce Type: new Abstract: In this paper, we advocate for incorporating enactive approaches to perception and cognition into artificial intelligence (AI). Enactive approaches view

Toward Reliable Design of LLM-Enabled Agentic Workflows: Optimizing Latency-Reliability-Cost Tradeoffs

SafetyDGX agent

arXiv:2605.23929v1 Announce Type: new Abstract: Modern AI systems increasingly rely on workflows composed of multiple interacting agents, some powered by large language models (LLMs) and others by con

Towards a Universal Causal Reasoner

ApplicationsDGX agent

arXiv:2605.24873v1 Announce Type: cross Abstract: Despite the importance of causal reasoning, training LLMs to reason causally remains underexplored. Existing data efforts mostly focus on benchmarking

Towards Cognitively-Faithful Decision-Making Models to Improve AI Alignment

SafetyDGX agent

arXiv:2509.04445v2 Announce Type: replace Abstract: Recent AI trends seek to align AI models to learned human-centric objectives, such as personal preferences, utility, or societal values. Using stand

Towards end-to-end LLM-based censoring-aware survival analysis

ResearchDGX agent

arXiv:2605.25399v1 Announce Type: new Abstract: Objective: Survival analysis is central to medical prediction, yet large language models (LLMs) are rarely used as end-to-end survival models because ce

Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild

ResearchDGX agent

arXiv:2605.24213v1 Announce Type: cross Abstract: Evaluation harnesses are software systems that orchestrate model evaluation by managing model invocation, data loading, metric computation, and result

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content

SafetyDGX agent

arXiv:2509.12672v2 Announce Type: replace Abstract: The volume of machine-generated content online has grown dramatically due to the widespread use of Large Language Models (LLMs), leading to new chal

Towards Large Model Feature Coding

Model ReleasesDGX agent

arXiv:2605.24025v1 Announce Type: cross Abstract: Large models have delivered remarkable performance across a wide range of perception and generation tasks, yet practical deployment is increasingly co

Towards Long-Horizon Interpretability: Efficient and Faithful Multi-Token Attribution for Reasoning LLMs

ResearchDGX agent

arXiv:2602.01914v2 Announce Type: replace Abstract: Token attribution methods provide intuitive explanations for language model outputs by identifying causally important input tokens. However, as mode

Towards Low-Gravity Planetary Exploration using Reinforcement Learning for Walking, Jumping, and In-flight Attitude Control

SafetyDGX agent

arXiv:2605.24643v1 Announce Type: new Abstract: This paper presents reinforcement learning (RL) policies for dynamic quadrupedal locomotion in planetary exploration scenarios. Building on a taskoptimi

← Previous
1…606607608609610…1040
Next →