AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
12 May 2026

Muown: Row-Norm Control for Muon Optimization

ResearchDGX agent

arXiv:2605.10797v1 Announce Type: new Abstract: Muon has emerged as a strong competitor to AdamW for language model pre-training, yet its behavior at scale is sensitive to weight decay. Recent work ha

MUR: Momentum Uncertainty guided Reasoning

TutorialsDGX agent

arXiv:2507.14958v2 Announce Type: replace Abstract: Current models have achieved impressive performance on reasoning-intensive tasks, yet optimizing their reasoning efficiency remains an open challeng

Nautilus Compass: Black-box Persona Drift Detection for Production LLM Agents

Model ReleasesDGX agent

arXiv:2605.09863v1 Announce Type: cross Abstract: Production LLM coding agents drift over long sessions: they forget user-specified constraints, slip into mistakes the user already flagged, and confab

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers

SafetyDGX agent

arXiv:2605.09176v1 Announce Type: cross Abstract: Training large language models requires optimization algorithms that are not only statistically effective, but also computationally and memory efficie

Neural Posterior Estimation of Terrain Parameters from Radar Sounder Data

Model ReleasesDGX agent

arXiv:2605.08179v1 Announce Type: cross Abstract: Radar sounders are electromagnetic instruments that can probe deep into the subsurface of Earth and other planetary bodies by processing the echo of t

Neuroscience-Inspired Analyses of Visual Interestingness in Multimodal Transformers

ResearchDGX agent

arXiv:2605.08188v1 Announce Type: cross Abstract: Human attention is the gateway to conscious perception, memory and decision-making. However, its role in modern transformer models remains largely une

NICE FACT: Diagnosing and Calibrating VLMs in Quantitative Reasoning for Kinematic Physics

TutorialsDGX agent

arXiv:2605.08452v1 Announce Type: new Abstract: The ability to derive precise spatial and physical insights is a cornerstone of vision-language models (VLMs), yet their poor performances in related sp

Octopus Protocol: One-Shot Hardware Discovery and Control for AI Agents via Infrastructure-as-Prompts

HardwareDGX agent

arXiv:2605.09055v1 Announce Type: cross Abstract: Recent agentic-robotics systems, from Code-asPolicies to modern vision-language-action (VLA) foundation models, presuppose that drivers, SDKs, or ROS-

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs

SafetyDGX agent

arXiv:2605.09433v1 Announce Type: new Abstract: Existing preference datasets for text-to-image models typically store only the final winner/loser images. This representation is insufficient for rectif

Optimality of Sub-network Laplace Approximations: New Results and Methods

Model ReleasesDGX agent

arXiv:2605.09075v1 Announce Type: cross Abstract: Although the Laplace approximation offers a simple route to uncertainty quantification in deep neural networks, its reliance on inverting large Hessia

Optimised Support Vector Regression for California Housing Price Prediction: The Critical Role of Feature Engineering and Hyperparameter Tuning

Model ReleasesDGX agent

arXiv:2605.08660v1 Announce Type: new Abstract: In the recent literature, Support Vector Regression (SVR) has been cited as one of the weakest performers on the California Housing benchmark dataset, w

OrderFusion: Encoding Orderbook for End-to-End Probabilistic Intraday Electricity Price Forecasting

Model ReleasesDGX agent

arXiv:2502.06830v5 Announce Type: replace-cross Abstract: Probabilistic intraday electricity price forecasting is becoming increasingly important for short-term power-system operation. With increasing

PACT: Peak-Aware Cross-Attention Graph Transformers for Efficient Storm-Surge Emulation

ResearchDGX agent

arXiv:2605.09036v1 Announce Type: new Abstract: Accurate and efficient storm-surge emulation is essential for coastal hazard assessment, yet high-fidelity hydrodynamic models remain too expensive for

Pairwise is Not Enough: Hypergraph Neural Networks for Multi-Agent Pathfinding

Model ReleasesDGX agent

arXiv:2602.06733v2 Announce Type: replace-cross Abstract: Multi-Agent Path Finding (MAPF) is a representative multi-agent coordination problem, where multiple agents are required to navigate to their

Parameterized Complexity of Stationarity Testing for Piecewise-Affine Functions and Shallow CNN Losses

Model ReleasesDGX agent

arXiv:2605.10219v1 Announce Type: cross Abstract: We study the parameterized complexity of testing approximate first-order stationarity at a prescribed point for continuous piecewise-affine (PA) funct

PDEAgent-Bench: A Multi-Metric, Multi-Library Benchmark for PDE Solver Generation

Model ReleasesDGX agent

arXiv:2605.09636v1 Announce Type: new Abstract: PDE-to-solver code generation aims to automatically synthesize executable numerical solvers from partial differential equation (PDE) specifications. Thi

Perceptual Asymmetry Between Hue Categories: Evidence from Human Color Categorization

ResearchDGX agent

arXiv:2605.09339v1 Announce Type: cross Abstract: Human color categories are not uniformly distributed in perceptual space, yet most computational color models still assume fixed and evenly structured

Plan2Cleanse: Test-Time Backdoor Defense via Monte-Carlo Planning in Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.09638v1 Announce Type: new Abstract: Ensuring the security of reinforcement learning (RL) models is critical, particularly when they are trained by third parties and deployed in real-world

PolarVSR: A Unified Framework and Benchmark for Continuous Space-Time Polarization Video Reconstruction

Model ReleasesDGX agent

arXiv:2605.10275v1 Announce Type: new Abstract: Polarimetric imaging captures surface polarization characteristics, such as the Degree of Linear Polarization (DoLP) and the Angle of Polarization (AoP)

Polygon-mamba: Retinal vessel segmentation using polygon scanning mamba and space-frequency collaborative attention

Local AiDGX agent

arXiv:2605.10581v1 Announce Type: new Abstract: Retinal vessel segmentation is crucial for diagnosis and assessment of ocular diseases. Notably, segmentation of small retinal vessels has been consiste

Position: Avoid Overstretching LLMs for every Enterprise Task

ApplicationsDGX agent

arXiv:2605.09365v1 Announce Type: new Abstract: Enterprise workloads are dominated by deterministic, structured, and knowledge-dependent tasks operating under strict cost, latency, and reliability con

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Model ReleasesDGX agent

arXiv:2605.08687v1 Announce Type: cross Abstract: Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user inter

Primal-Dual Guided Decoding for Constrained Discrete Diffusion

SafetyDGX agent

arXiv:2605.09749v1 Announce Type: new Abstract: Discrete diffusion models generate structured sequences by progressively unmasking tokens, but enforcing global property constraints during generation r

PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines

Model ReleasesDGX agent

arXiv:2605.10614v1 Announce Type: new Abstract: Multi-agent LLM systems introduce a security risk in which sensitive information accessed by one agent can propagate through shared context and reappear

Prompt Estimation from Prototypes for Federated Prompt Tuning of Vision Transformers

Model ReleasesDGX agent

arXiv:2510.25372v2 Announce Type: replace Abstract: Visual Prompt Tuning (VPT) of pre-trained Vision Transformers (ViTs) has proven highly effective as a parameter-efficient fine-tuning technique for

PumpSense: Real-Time Detection and Target Extraction of Crypto Pump-and-Dumps on Telegram

Model ReleasesDGX agent

arXiv:2605.09431v1 Announce Type: new Abstract: Cryptocurrency pump-and-dump schemes coordinated via Telegram threaten market integrity. However, existing research addressing this specific threat has

Quantifying the Utility of User Simulators for Building Collaborative LLM Assistants

Model ReleasesDGX agent

arXiv:2605.09808v1 Announce Type: new Abstract: User simulators are increasingly leveraged to build interactive AI assistants, yet how to measure the quality of these simulators remains an open questi

Quantile-Coupled Flow Matching for Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2605.08515v1 Announce Type: new Abstract: Unlike standard expected-return Reinforcement Learning (RL), Distributional RL (DRL) models the full return distribution, making it better-suited for un

Quantitative Sobolev Approximation Bounds for Neural Operators with Empirical Validation on Burgers Equation

Model ReleasesDGX agent

arXiv:2605.08170v1 Announce Type: new Abstract: Neural operators have emerged as a powerful tool for learning mappings between infinite-dimensional function spaces. However, their approximation proper

Queryable LoRA: Instruction-Regularized Routing Over Shared Low-Rank Update Atoms

Model ReleasesDGX agent

arXiv:2605.08423v1 Announce Type: cross Abstract: We present a data-adaptive method for parameter-efficient fine-tuning of large neural networks. Standard low-rank adaptation methods improve efficienc

R4Det: 4D Radar-Camera Fusion for High-Performance 3D Object Detection

Model ReleasesDGX agent

arXiv:2603.11566v2 Announce Type: replace Abstract: 4D radar-camera sensing configuration has gained increasing importance in autonomous driving. However, existing 3D object detection methods that fus

RareCP: Regime-Aware Retrieval for Efficient Conformal Prediction

Model ReleasesDGX agent

arXiv:2605.08857v1 Announce Type: new Abstract: Recent advances in uncertainty quantification for time series forecasting show that conformal prediction can provide reliable prediction intervals, yet

Recovering Physical Dynamics from Discrete Observations via Intrinsic Differential Consistency

Model ReleasesDGX agent

arXiv:2605.08454v1 Announce Type: cross Abstract: Recovering continuous-time dynamics from discrete observations is difficult because local supervision (e.g., pointwise regression targets, derivative

Relations Are Channels: Knowledge Graph Embedding via Kraus Decompositions

SafetyDGX agent

arXiv:2605.10317v1 Announce Type: cross Abstract: Knowledge graph embedding (KGE) models typically represent each relation as an operator on entity embeddings. In this work, we identify three structur

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing

ResearchDGX agent

arXiv:2605.08840v1 Announce Type: new Abstract: Large language models (LLMs) face growing challenges in efficient generative inference due to the increasing memory demands of Key-Value (KV) caches, es

Rethinking Random Transformers as Adaptive Sequence Smoothers for Sleep Staging

Model ReleasesDGX agent

arXiv:2605.09905v1 Announce Type: cross Abstract: Automatic sleep staging commonly adopts Transformers under the assumption that they learn complex long-range dependencies. We challenge this view by r

Robust Spectral Watermark for Synthetic Tabular Data

Model ReleasesDGX agent

arXiv:2511.21600v2 Announce Type: replace-cross Abstract: The rise of generative AI has enabled the production of high-fidelity synthetic tabular data across fields such as healthcare, finance, and pu

RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild

Model ReleasesDGX agent

arXiv:2605.10357v1 Announce Type: cross Abstract: Multimodal misinformation increasingly leverages visual persuasion, where repurposed or manipulated images strengthen misleading text. We introduce ex

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.08391v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning agents that act on partial local observations face a fundamental information bottleneck: the knowledge ne

SAP SAPPHIRE 2026: Google Cloud unveils unified agentic vision and massive compute scaling

Model ReleasesDGX agent

In today's hyper-connected market, an enterprise's most valuable asset — mission-critical data — often remains trapped in legacy silos. For years, leadership teams have navigated a data pipeline dilem

SearchSkill: Teaching LLMs to Use Search Tools with Evolving Skill Banks

ResearchDGX agent

arXiv:2605.09038v1 Announce Type: new Abstract: Teaching language models to use search tools is not only a question of whether they search, but also of whether they issue good queries. This is especia

Selective LoRA for Visual Tokens and Attention Heads

Model ReleasesDGX agent

arXiv:2512.19219v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is widely used for parameter-efficient fine-tuning, but its standard all-token, all-head design ignores the heterog

Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation

SafetyDGX agent

arXiv:2605.08874v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation requires adapting image-level vision-language models such as CLIP to dense pixel-level prediction, which is challe

Shaping Schema via Language Representation as the Next Frontier for LLM Intelligence Expanding

ApplicationsDGX agent

arXiv:2605.09271v1 Announce Type: new Abstract: Although natural language is the default medium for Large Language Models (LLMs), its limited expressive capacity creates a profound bottleneck for comp

Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success

Model ReleasesDGX agent

arXiv:2605.09070v1 Announce Type: cross Abstract: Many jailbreak attack research papers report attack success rates for a limited number of parameter settings, even though there are many combinations

Sketch-and-Verify: Structured Inference-Time Scaling via Program Sketching

Model ReleasesDGX agent

arXiv:2605.08658v1 Announce Type: cross Abstract: SKETCHVERIFY is a within-tier cost-performance policy, not a universal accuracy improvement. The operational question: a practitioner stuck with a sma

Sources: Anthropic is in talks to raise between 30B and 50B in a funding round that would value it at up to $950B (Mike Isaac/New York Times)

Model ReleasesDGX agent

Mike Isaac / New York Times: Sources: Anthropic is in talks to raise between 30B and 50B in a funding round that would value it at up to 950B — The start-up, which recently released a powerful A.I. mo

Sparsity Moves Computation: How FFN Architecture Reshapes Attention in Small Transformers

Model ReleasesDGX agent

arXiv:2605.09403v1 Announce Type: cross Abstract: Architectural choices inside the Transformer feedforward network (FFN) block do not merely affect the block itself; they reshape the computations lear

Spatial Priming Outperforms Semantic Prompting: A Grid-Based Approach to Improving LLM Accuracy on Chart Data Extraction

ResearchDGX agent

arXiv:2605.08220v1 Announce Type: new Abstract: The automated extraction of data from scientific charts is a critical task for large-scale literature analysis. While multimodal Large Language Models (

Spectral Condition for muP under Width-Depth Scaling

ResearchDGX agent

arXiv:2603.00541v2 Announce Type: replace Abstract: Generative foundation models are increasingly scaled in both width and depth, posing significant challenges for stable feature learning and reliable

Step Rejection Fine-Tuning: A Practical Distillation Recipe

Model ReleasesDGX agent

arXiv:2605.10674v1 Announce Type: cross Abstract: Rejection Fine-Tuning (RFT) is a standard method for training LLM agents, where unsuccessful trajectories are discarded from the training set. In the

Strategic commitments shape collective cybersecurity under AI inequality

Model ReleasesDGX agent

arXiv:2605.09415v1 Announce Type: new Abstract: The growing integration of AI into cybersecurity is reshaping the balance between attackers and defenders. When access to advanced AI-enabled defence to

TAH-QUANT: Effective Activation Quantization in Pipeline Parallelism over Slow Network

ResearchDGX agent

arXiv:2506.01352v2 Announce Type: replace Abstract: Decentralized training of large language models offers the opportunity to pool computational resources across geographically distributed participant

Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation

Model ReleasesDGX agent

arXiv:2505.11604v5 Announce Type: replace Abstract: Editing presentation slides is a frequent yet tedious task, ranging from creative layout design to repetitive text maintenance. While recent GUI-bas

Task complexity shapes internal representations and robustness in neural networks

ResearchDGX agent

arXiv:2508.05463v2 Announce Type: replace-cross Abstract: Neural networks excel across a wide range of tasks, yet remain black boxes. In particular, how their internal representations are shaped by th

The autoPET3 Challenge: Automated Lesion Segmentation in Whole-Body PET/CT nicode{x2013} Multitracer Multicenter Generalization

Model ReleasesDGX agent

arXiv:2605.05775v2 Announce Type: replace-cross Abstract: We report the design and results of the third autoPET challenge (MICCAI 2024), which benchmarked automated lesion segmentation in whole-body P

The Bystander Effect in Multi-Agent Reasoning: Quantifying Cognitive Loafing in Collaborative Interactions

SafetyDGX agent

arXiv:2605.10698v1 Announce Type: cross Abstract: Multi-agent systems (MAS) assume that collaborating inherently improves Large Language Model (LLM) reasoning. We challenge this by demonstrating that

The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies

Model ReleasesDGX agent

arXiv:2605.10799v1 Announce Type: cross Abstract: Corruption studies, the primary tool for evaluating chain-of-thought (CoT) faithfulness, identify which chain positions are 'computationally important

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using…

Model ReleasesDGX agent

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using AI in assistant mode to study can hurt learning because it

Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving

AgentsDGX agent

arXiv:2605.10117v1 Announce Type: cross Abstract: Autonomous driving scenes range from empty highways to dense intersections with dozens of interacting road users, yet current 3D detection models appl

← Previous
1…562563564565566…1061
Next →