AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
26 May 2026

SeqRoute: Global Budget-Aware Sequential LLM Routing via Offline Reinforcement Learning

ApplicationsDGX agent

arXiv:2605.25424v1 Announce Type: cross Abstract: Existing LLM routing frameworks treat queries as independent events, neglecting the sequential nature of real-world user sessions constrained by globa

Side-by-side Comparison Amplifies Dialect Bias in Language Models

SafetyDGX agent

arXiv:2605.24384v1 Announce Type: cross Abstract: Language models (LMs) can exhibit systematic biases against speakers based on variations in their dialects, even in the absence of a dialect label, a

Signs Beat Floats: Low-Rank Double-Binary Adaptation for On-Device Fine-Tuning

Local AiDGX agent

arXiv:2605.24058v1 Announce Type: cross Abstract: On-device adaptation of large language models commonly keeps a quantized base model frozen while training and deploying a small, task-specific LoRA ad

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Simulating Human Memory with Language Models

ApplicationsDGX agent

arXiv:2605.25680v1 Announce Type: cross Abstract: Language models are increasingly being deployed as user simulators, but their memory is far more reliable than that of real users. To measure this gap

'Si'multaneous 'S'patial-'T'emporal Message Passing for Dynamic Graph Representation Learning

Model ReleasesDGX agent

arXiv:2605.25548v1 Announce Type: cross Abstract: Dynamic graph neural networks (DGNNs) that operate on snapshot sequences typically fall into one of two categories. Temporal-first approaches build pe

SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking

Model ReleasesDGX agent

arXiv:2605.25160v1 Announce Type: new Abstract: Mobile GUI agents powered by large language models have progressed rapidly, creating urgent needs for realistic and comprehensive evaluation. Existing b

Single View Seafloor Recovery from Imaging Sonar via Differentiable Rendering

ResearchDGX agent

arXiv:2605.24195v1 Announce Type: cross Abstract: Sonar is often the only modality suitable for high-resolution imaging underwater due to light attenuation and turbidity. Forward-looking imaging sonar

SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills

Model ReleasesDGX agent

arXiv:2605.24117v1 Announce Type: new Abstract: Large language model (LLM) agents accumulate rich episodic trajectories while solving real-world tasks, but it remains unclear whether such experience c

SLAP: Stratified Loss-based Pruning for On-Policy Data-Efficient Instruction Tuning

Model ReleasesDGX agent

arXiv:2605.23969v1 Announce Type: new Abstract: Instruction tuning has optimized the specialized capabilities of large language models (LLMs), but it often requires extensive datasets and prolonged tr

SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation

SafetyDGX agent

arXiv:2605.24371v1 Announce Type: cross Abstract: CT report generation (CTRG) requires models to summarize three-dimensional anatomical context and pathological findings from hundreds of axial slices.

Small Ensemble-based Data Assimilation: A Machine Learning-Enhanced Data Assimilation Method with Limited Ensemble Size

TutorialsDGX agent

arXiv:2510.15284v2 Announce Type: replace Abstract: Ensemble-based data assimilation (DA) methods have become increasingly popular due to their inherent ability to address nonlinear dynamic problems.

Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural PDE Solvers

Model ReleasesDGX agent

arXiv:2605.25949v1 Announce Type: cross Abstract: Neural PDE solvers have followed the scaling trajectory of vision and language, with recent foundation models reaching billions of parameters. We argu

Smart Timing for Mining: A Deep Learning Framework for Bitcoin Hardware ROI Prediction

Model ReleasesDGX agent

arXiv:2512.05402v2 Announce Type: replace-cross Abstract: Bitcoin mining hardware acquisition requires strategic timing due to volatile markets, rapid technological obsolescence, and protocol-driven r

SMDD-Bench: Can LLMs Solve Real-World Small Molecule Drug Design Tasks?

Model ReleasesDGX agent

arXiv:2605.21740v2 Announce Type: replace Abstract: LLM agents have incredible potential for scientific discovery applications. However, the performance of LLM agents on real-world, small molecule dru

Smoother Action Chunking Flow Policy via Prior-Corrected Orthogonal Trust-Region Guidance

SafetyDGX agent

arXiv:2605.24433v1 Announce Type: cross Abstract: Flow-matching robot policies commonly use action-chunking inference for efficient closed-loop control, but chunk boundaries can introduce discontinuou

SODE: Analyzing Social Dynamics in LLM Agents

Model ReleasesDGX agent

arXiv:2605.23949v1 Announce Type: cross Abstract: As Large Language Models (LLMs) evolve into interactive agents, understanding their behavioral alignment within human social dynamics becomes essentia

Soft Pneumatic Actuators for Soft Robotics: A Motion-Based Review of Actuation Mechanisms and Performance Trade-offs

TutorialsDGX agent

arXiv:2605.25109v1 Announce Type: new Abstract: Soft pneumatic actuators are widely used in soft robotics because they can produce large motions while remaining compliant enough to interact safely wit

Soft Pneumatic Grippers: Topology optimization, 3D-printing and Experimental validation

ResearchDGX agent

arXiv:2511.19211v3 Announce Type: replace Abstract: This paper presents a systematic topology optimization framework for designing a soft pneumatic gripper (SPG), explicitly considering the design-dep

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models

Model ReleasesDGX agent

arXiv:2506.18543v2 Announce Type: replace-cross Abstract: The rapid proliferation of Large Language Models (LLMs) has heightened concerns regarding their exposure to jailbreak attacks, which craft adv

SoK: DARPA's AI Cyber Challenge (AIxCC): Competition Design, Architectures, and Lessons Learned

AgentsDGX agent

arXiv:2602.07666v3 Announce Type: replace-cross Abstract: DARPA's AI Cyber Challenge (AIxCC, 2023--2025) is the largest competition to date for building fully autonomous cyber reasoning systems (CRSs)

Solving Combinatorial Counting Problems with Weighted First-Order Model Counting

ResearchDGX agent

arXiv:2605.24845v1 Announce Type: new Abstract: Combinatorial counting problems pervade artificial intelligence, statistics, and discrete mathematics. Whether the task is enumerating subsets, multiset

SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models

Model ReleasesDGX agent

arXiv:2605.25420v1 Announce Type: cross Abstract: Large language model safety evaluation remains heavily English-centered, leaving low-resource languages under-measured even when models are deployed g

Some Robustness Properties of Label Cleaning

ResearchDGX agent

arXiv:2509.11379v3 Announce Type: replace-cross Abstract: We demonstrate that learning procedures that rely on aggregated labels, e.g., label information distilled from noisy responses, enjoy robustne

SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models

ResearchDGX agent

arXiv:2602.02544v2 Announce Type: replace-cross Abstract: While Diffusion Language Models (DLMs) offer a flexible, arbitrary-order alternative to the autoregressive paradigm, their non-causal nature p

SPACE: Unifying Symmetric and Asymmetric Routing Problems for Generalist Neural Solver

ApplicationsDGX agent

arXiv:2605.24484v1 Announce Type: new Abstract: Generalist neural routing solvers have shown great potential in solving diverse vehicle routing problems (VRPs) with a unified model. However, existing

Spacetime Formation under Requirements: Contextual Realization and Form-Dependent Probability

ResearchDGX agent

arXiv:2605.23943v1 Announce Type: new Abstract: Quantum cognition often explains order effects, contextuality, and violations of the law of total probability by replacing classical probability with qu

SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharing

AgentsDGX agent

arXiv:2512.24008v3 Announce Type: replace Abstract: Personalized search demands the ability to model users' evolving, multi-dimensional information needs; a challenge for systems constrained by static

Spatio-temporal, multi-field deep learning of shock propagation in meso-structured media

Local AiDGX agent

arXiv:2509.16139v5 Announce Type: replace Abstract: Predicting the extreme hydrodynamic response of porous and architected lattice materials is a fundamental challenge in high energy density physics,

SpecAlign: A Semantic Alignment Framework for SystemVerilog Assertion Generation

SafetyDGX agent

arXiv:2605.25181v1 Announce Type: new Abstract: Existing Large Language Model (LLM) approaches to SystemVerilog Assertion (SVA) generation primarily focus on syntactic validity and formal verification

Specification-Based Code-Text-Code Reengineering for LLM-Mediated Software Evolution

ResearchDGX agent

arXiv:2605.25232v1 Announce Type: cross Abstract: Direct Code2Code transformation remains challenging to control because it can preserve surface-level syntax while introducing semantic drift, hidden b

SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning

Local AiDGX agent

arXiv:2509.05614v3 Announce Type: replace-cross Abstract: Pruning is a typical acceleration technique for compute-bound models by removing computation on unimportant values. Recently, it has been appl

Spectral Probe-Circuits: A Three-Step Recipe for Identifying Attention-Head Circuits in Pretrained Transformers

Model ReleasesDGX agent

arXiv:2605.24059v1 Announce Type: cross Abstract: We present a three-step recipe for identifying attention-head circuits in pretrained transformers. A per-head spectral signal -- the time-integrated p

Spectral Retrieval: Multi-Scale Sinc Convolution over Token Embeddings for Localized Retrieval in LLM Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.24764v1 Announce Type: cross Abstract: [Abridged] - Spectral Retrieval is a plug-in re-ranking stage that interpolates between per-token MaxSim and mean-pool retrieval through a multi-scale

Spiking the training data to correct for test set contamination

ResearchDGX agent

arXiv:2605.24818v1 Announce Type: cross Abstract: The literature on test set contamination largely focuses on detection, but the correction of contaminated test scores is underexplored. Our core propo

Split-Merge: A Difference-based Approach for Dominant Eigenvalue Problem

Model ReleasesDGX agent

arXiv:2501.15131v3 Announce Type: replace-cross Abstract: The computation of the dominant eigenpair for symmetric positive semidefinite matrices is fundamental in numerical optimization. This work shi

Spurious Stationarity and Hardness Results for Bregman Proximal-Type Algorithms

ResearchDGX agent

arXiv:2404.08073v3 Announce Type: replace-cross Abstract: Bregman proximal-type algorithms (BPs), such as mirror descent, have become popular tools in machine learning and data science for exploiting

Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation

ResearchDGX agent

arXiv:2605.26111v1 Announce Type: cross Abstract: Subject-driven image generation aims to synthesize new images that preserve the identity of the given subject while following textual instructions. Ex

StakeBench: Evaluating Language Understanding Grounded in Market Commitment

SafetyDGX agent

arXiv:2605.26074v1 Announce Type: cross Abstract: Existing financial NLP benchmarks often rely on labels supplied by outside observers, measuring how language is perceived rather than what speakers ha

STaT: Resolving Shape Distortion in Non-Stationary Time Series via Tri-Modal Synergy

SafetyDGX agent

arXiv:2605.25943v1 Announce Type: new Abstract: Recent research in time series forecasting frequently investigates the integration of textual and visual modalities with numerical models to better navi

Statistical Inference for Stochastic Gradient Descent Beyond Finite Variance

ResearchDGX agent

arXiv:2605.26000v1 Announce Type: cross Abstract: Stochastic gradient descent (SGD) is a foundational algorithm for large-scale statistical learning and stochastic optimization. However, statistical i

Steering Beyond the Support: Adversarial Training on Unsupervised Jailbroken Activation Simulation

SafetyDGX agent

arXiv:2605.24535v1 Announce Type: cross Abstract: Jailbreak prompts can trigger harmful completions on aligned LLMs, In accordance, safety steering has been proposed: test-time activation intervention

Stein Variational Ergodic Surface Coverage with SE(3) Constraints

Model ReleasesDGX agent

arXiv:2603.09458v3 Announce Type: replace Abstract: Surface manipulation tasks require robots to generate trajectories that comprehensively cover complex 3D surfaces while maintaining precise end-effe

Step-TP: A Grounded, Step-Level Dataset with Chain-of-Thought Reasoning for LLM-Guided Tensor Program Optimization

ResearchDGX agent

arXiv:2605.25954v1 Announce Type: cross Abstract: Despite the strong reasoning capabilities of large language models (LLMs), optimizing the execution efficiency of tensor programs remains challenging

StepGap: A Hybrid NLI-LLM Checker for Step-Level Evidence-Gap Detectionin Multi-Hop Question Answering

ResearchDGX agent

arXiv:2605.24733v1 Announce Type: new Abstract: We present extbf{StepGap}, a hybrid NLI-LLM decision tree that detects step-level evidence gaps in multi-hop QA and emits one of three typed labels: ext

Stiffness Optimization for Concentrated Bending in Magnetically Actuated Catheters: Maintaining Steerability under Gradient Stiffness

ResearchDGX agent

arXiv:2605.25005v1 Announce Type: new Abstract: Achieving both efficient pushability (propulsion transmission) and proximally concentrated bending for steerability is challenging for magnetically actu

Stochastic Estimation of the Layer-wise Hessian Trace for Monitoring Neural-network Training

Model ReleasesDGX agent

arXiv:2605.25674v1 Announce Type: new Abstract: The loss and the norm of its gradient separate the healthy and the pathological regimes of neural-network training only weakly, whilst the curvature of

Stochastic Linear Bandits with Parameter Noise

Model ReleasesDGX agent

arXiv:2601.23164v2 Announce Type: replace Abstract: We study the stochastic linear bandits with parameter noise model, in which the reward of action a is a^op heta where heta is sampled i.i.d. We show

Stop Comparing LLM Agents Without Disclosing the Harness

SafetyDGX agent

arXiv:2605.23950v1 Announce Type: new Abstract: This position paper argues that, for long-horizon tasks evaluated across models with comparable frontier capability, the agent execution harness, namely

STORM: Internalized Modeling for Spatial-Temporal Reasoning in Video-Language Models

AgentsDGX agent

arXiv:2605.26014v1 Announce Type: cross Abstract: Many video reasoning tasks require tracking motion, temporal order, and evolving visual states across frames. Existing methods built on large vision-l

Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games

SafetyDGX agent

arXiv:2605.04906v2 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint

STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media

Model ReleasesDGX agent

arXiv:2605.25162v1 Announce Type: cross Abstract: Large language models for vertical domains are bottlenecked by the scarcity of complex, domain-specific task-oriented dialogues. Existing data acquisi

Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning

Model ReleasesDGX agent

arXiv:2605.24709v1 Announce Type: new Abstract: Streaming reinforcement learning has emerged as an online learning paradigm that conforms to the restrictions of natural learning agents that process da

StreamProfileBench: A Benchmark for Fine-Grained User Profile Inference in Real-World Streaming Scenarios

Model ReleasesDGX agent

arXiv:2605.25758v1 Announce Type: new Abstract: Large Language Models (LLMs) have reshaped user profiling, yet current evaluations mainly focus on static data snapshots. This paradigm overlooks the re

StrTransformer: Source-Wise Structured Transformers for Unsupervised Blind Source Recovery

ApplicationsDGX agent

arXiv:2605.25648v1 Announce Type: cross Abstract: This paper proposes StrTransformer, a source-wise structured Transformer framework for blind source recovery and branch-wise latent modeling. Instead

StructBreak: Structural Cognitive Overload-Induced Safety Failures in MLLMs

Model ReleasesDGX agent

arXiv:2605.25534v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at structural reasoning yet suffer from a sharp logical brittleness in structural consistency. We term th

Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training

Model ReleasesDGX agent

arXiv:2603.17198v2 Announce Type: replace-cross Abstract: A foundational principle in cognitive science holds that intelligent agents do not learn by storing experiences as isolated instances, but by

Structure-Aware RAG: Structured Retrieval Augmented Generation from Noisy Data for Conversational Agents

ApplicationsDGX agent

arXiv:2605.24366v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely adopted in conversational applications. However, their reliance on parametric knowledge limits reliability

Subspace Aggregation Query and Index Generation for Multidimensional Resource Space Model

ResearchDGX agent

arXiv:2505.02129v3 Announce Type: replace-cross Abstract: Organizing large-scale resources in a multidimensional semantic space is an approach to efficiently managing and querying resources from diffe

Subspace-Guided Semantic and Topological Invariant Registration for Annotation-Free Ultrasound Plane Quality Control

SafetyDGX agent

arXiv:2605.25396v1 Announce Type: cross Abstract: Reliable quality control (QC) of ultrasound images is essential for both real-time acquisition guidance and retrospective clinical audit, yet existing

Sum of Costs Diffusion with Dynamic Guidance for Motion Planning

ResearchDGX agent

arXiv:2605.24690v1 Announce Type: cross Abstract: The motion planning problem for robotic manipulation can be addressed through classical or deep learning approaches. Existing methods face significant

← Previous
1…605606607608609…1040
Next →