AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
2 Jun 2026

How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval

AgentsDGX agent

arXiv:2606.00308v1 Announce Type: cross Abstract: Large-language-model code generation has shifted from single-shot prompting to multi-agent orchestrations - analyst, coder, tester, and debugger pipel

Hyperspherical Variational Autoencoders Using Efficient Spherical Cauchy Distribution

Model ReleasesDGX agent

arXiv:2506.21278v3 Announce Type: replace-cross Abstract: We propose spherical Cauchy (spCauchy) latent variables for variational autoencoders on hyperspherical latent spaces. The spCauchy family has

hZACH-ViT: Curved Latent Geometry for Compact Vision Transformers in Low-Data Medical Imaging

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.00906v1 Announce Type: new Abstract: Compact Vision Transformers are attractive for medical imaging in low-data and resource-constrained settings, but most existing variants assume that Euc

IMAC-AgriVLN: Can Agricultural Vision-and-Language Navigation Agents be Aware of Instruction Mistakes?

Model ReleasesDGX agent

arXiv:2606.02519v1 Announce Type: new Abstract: Agricultural robots are serving as powerful assistants across a wide range of agricultural tasks, nevertheless, still heavily relying on manual operatio

InstructSAM: Segment Any Instance with Any Instructions

Model ReleasesDGX agent

arXiv:2605.26102v2 Announce Type: replace Abstract: In this paper, we introduce InstructSAM, a unified and streamlined framework designed for multi-instance segmentation under arbitrary instructions.

Joint Agent Memory and Exploration Learning via Novelty Signals

SafetyDGX agent

arXiv:2606.01528v1 Announce Type: new Abstract: In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploratio

Latent Reward Steering: An Adaptive Inference-Time Framework that Implicitly Promotes Cognitive Behaviors in Reasoning LLMs

ResearchDGX agent

arXiv:2606.00726v1 Announce Type: new Abstract: Strong reasoning depends not only on model knowledge but also on how effectively cognitive behaviors are deployed during generation. Existing methods of

Law professors wrote questions they were asked during office hours. Gemini 2.5 & humans answered them then other law professors blindly judg…

Model ReleasesDGX agent

Law professors wrote questions they were asked during office hours. Gemini 2.5 & humans answered them then other law professors blindly judged the results: -Gemini had a 75% win rate vs. professors -G

Learning Chaotic Dynamics through Second-Order Geometric Supervision

Local AiDGX agent

arXiv:2606.01596v1 Announce Type: cross Abstract: Learning chaotic dynamical systems from data requires more than short-term predictive accuracy: the learned model must preserve the attractor geometry

Learning to Retrieve: Dual-Level Long-Term Memory for Text-to-SQL Agents

Model ReleasesDGX agent

arXiv:2606.00547v1 Announce Type: new Abstract: Interactive text-to-SQL agents solve database tasks through multi-turn interactions involving schema exploration, query execution, feedback interpretati

Learning When to Translate for Multilingual Reasoning

TutorialsDGX agent

arXiv:2606.02465v1 Announce Type: cross Abstract: Reasoning language models (RLMs) achieve strong performance on complex reasoning tasks, but still exhibit substantial multilingual reasoning gaps, lar

Linear Strategic Classification with Endogenous Improvements

ApplicationsDGX agent

arXiv:2606.01198v1 Announce Type: new Abstract: Strategic classification studies settings in which agents respond to a deployed classifier by modifying observable features at a cost. Classical models

MASER: Modality-Adaptive Specialist Routing for Embodied 3D Spatial Intelligence

Model ReleasesDGX agent

arXiv:2606.02463v1 Announce Type: cross Abstract: In 3D environments, Embodied Agents answer spatially relevant questions through reasoning from a mixture of modalities including natural language, RGB

MAVL: A Multilingual Audio-Video Lyrics Dataset for Animated Song Translation

Model ReleasesDGX agent

arXiv:2505.18614v5 Announce Type: replace Abstract: Lyrics translation requires both accurate semantic transfer and preservation of musical rhythm, syllabic structure, and poetic style. In animated mu

MidSurfNet: Learnable Face Pairing and Interference Implicit Fields for Generalized Mid-surface Abstraction

ApplicationsDGX agent

arXiv:2606.01891v1 Announce Type: cross Abstract: Mid-surface abstraction is essential for finite element analysis of thin-walled CAD models. Existing face pairing-based methods rely on handcrafted ge

Molecular Embedding-Based Algorithm Selection in Protein-Ligand Docking

Model ReleasesDGX agent

arXiv:2512.02328v2 Announce Type: replace-cross Abstract: Selecting an effective docking algorithm is highly context-dependent, and no single method performs reliably across structural, chemical, and

MOSS-Audio Technical Report

ResearchDGX agent

arXiv:2606.01802v1 Announce Type: cross Abstract: MOSS-Audio is a unified audio-language model for speech, environmental sound, and music understanding, supporting audio captioning, time-aware questio

Motion-aware Event Suppression for Event Cameras

Model ReleasesDGX agent

arXiv:2602.23204v3 Announce Type: replace Abstract: In this work, we introduce the first framework for Motion-aware Event Suppression, which learns to filter events triggered by IMOs and ego-motion in

Multi-Agent Computer Use

Model ReleasesDGX agent

arXiv:2606.01533v1 Announce Type: cross Abstract: Computer use agents (CUAs) today are primarily deployed as single serial agents. This setup is suboptimal for complex long-horizon tasks that benefit

naPINN: Noise-Adaptive Physics-Informed Neural Networks for Recovering Physics from Corrupted Measurement

Model ReleasesDGX agent

arXiv:2602.02547v2 Announce Type: replace-cross Abstract: Physics-Informed Neural Networks (PINNs) are effective methods for solving inverse problems and discovering governing equations from observati

Near-Optimal Pure Machine Unlearning for Smooth Strongly Convex Losses

ApplicationsDGX agent

arXiv:2606.01527v1 Announce Type: new Abstract: Machine unlearning is motivated by legal and user-facing requirements to remove the influence of individuals' data from trained models, such as the righ

NILC: Discovering New Intents with LLM-assisted Clustering

Model ReleasesDGX agent

arXiv:2511.05913v2 Announce Type: replace-cross Abstract: New intent discovery (NID) seeks to recognize both new and known intents from unlabeled user utterances, which finds prevalent use in practica

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents

Model ReleasesDGX agent

arXiv:2606.02031v1 Announce Type: cross Abstract: Building capable visual web agents requires long-horizon reasoning, precise grounding, and robust interaction with dynamic real-world websites. Despit

Optimizing accuracy and diversity: a multi-task approach to forecast combinations

ResearchDGX agent

arXiv:2310.20545v3 Announce Type: replace Abstract: We present a multi-task optimization approach based on a deep learning architecture for time series forecasting. We leverage large collections of ti

PathAR: Structure-First Autoregressive Synthesis of Multimodal Pathology Images

ResearchDGX agent

arXiv:2606.01543v1 Announce Type: new Abstract: Data scarcity in multimodal pathology motivates unified generative models that synthesize modality-specific appearance while preserving anatomically coh

Perspective on Bias in Biomedical AI: Preventing Downstream Healthcare Disparities

SafetyDGX agent

arXiv:2604.14514v2 Announce Type: replace Abstract: Healthcare disparities persist across socioeconomic boundaries, often attributed to unequal access to screening, diagnostics, and therapeutics. Howe

PFT: Phonon Fine-tuning for Machine Learned Interatomic Potentials

Model ReleasesDGX agent

arXiv:2601.07742v4 Announce Type: replace-cross Abstract: Many materials properties depend on higher-order derivatives of the potential energy surface, yet machine learned interatomic potentials (MLIP

Pre-Deployment Robustness Stress Testing for CT Segmentation Systems Using Clinically Motivated Multi-Corruption Augmentation

Model ReleasesDGX agent

arXiv:2606.00491v1 Announce Type: cross Abstract: Deep learning-based CT segmentation systems often achieve high accuracy on clean benchmark images, but their performance may degrade under heterogeneo

Predicted-Flow Control Barrier Functions for Real-Time Safe Optimal Control

Model ReleasesDGX agent

arXiv:2606.00297v1 Announce Type: cross Abstract: Control barrier functions (CBFs) provide real-time safety guarantees through pointwise conditions on the state. However, synthesizing a valid CBF is d

Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation

SafetyDGX agent

arXiv:2602.07298v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) represent a promising frontier for recommender systems, yet their development has been impeded by the absence of

PrivacyPeek: Auditing What LLM-Based Agents Acquire, Not Just What They Say

Model ReleasesDGX agent

arXiv:2606.00152v1 Announce Type: cross Abstract: LLM-based agents are rapidly advancing, autonomously invoking external tools to complete multi-step tasks for users. However, agents often acquire mor

Quality-Guided Semi-Supervised Learning for Medical Image Segmentation

ResearchDGX agent

arXiv:2606.01753v1 Announce Type: new Abstract: Training accurate medical image segmentation models requires large amounts of densely annotated data, which is costly and time-consuming to obtain. Semi

Relative Energy Learning for LiDAR Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2511.06720v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection is a critical requirement for reliable autonomous driving, where safety depends on recognizing road obstacles an

RenoBench: A Citation Parsing Benchmark

Model ReleasesDGX agent

arXiv:2603.25640v2 Announce Type: replace-cross Abstract: Accurate parsing of citations is necessary for machine-readable scholarly infrastructure. But, despite sustained interest in this problem, exi

Reusing Fusion-Time Spectral Reliability for Adaptive Fusion and Expert Routing in RGB-Infrared Object Detection

Model ReleasesDGX agent

arXiv:2606.01173v1 Announce Type: new Abstract: RGB-infrared detectors typically discard the statistics generated during cross-modal fusion, leaving downstream modules unaware of whether the current i

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation

SafetyDGX agent

arXiv:2507.02792v5 Announce Type: replace Abstract: Text-to-image (T2I) diffusion models have shown remarkable success in generating high-quality images from text prompts. Recent efforts extend these

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment

SafetyDGX agent

arXiv:2606.02530v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) with human values often degrades their general capabilities, termed the alignment tax. Existing methods mitigate t

Self-Evolving Hermes Agents: Enterprise AI That Gets Better With Use | Nemotron Labs https://x.com/i/broadcasts/1pJdRRyneOjKW

Model ReleasesDGX agent

This likely describes a framework or system for deploying AI agents that autonomously improve their performance over time through continuous learning and adaptation in enterprise environments. The sel

Semi-Supervised Hyperbolic Hierarchical Clustering with Set-Level Structural Priors

Model ReleasesDGX agent

arXiv:2606.01525v1 Announce Type: new Abstract: Semi-supervised hierarchical clustering aims to learn a tree structure consistent with data patterns and user-provided supervision. Supervision is usual

Serving MiniMax-M3 for efficient inference: Unlocking 1M-Token Context and Multimodality Without Regrets

ToolsDGX agent

MiniMax-M3 is a large language model capable of handling 1 million token contexts and multimodal inputs while maintaining efficient inference performance. Together AI's blog post discusses techniques

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

Model ReleasesDGX agent

arXiv:2606.02380v1 Announce Type: cross Abstract: As LLM-based agents expand their operational scope, reliability becomes a prerequisite for real-world deployment. However, in practical applications,

Spatiotemporal Multi-Task Graph Transformer for Trip-Level Transit Prediction

SafetyDGX agent

arXiv:2606.00572v1 Announce Type: new Abstract: Passenger count data from public transit systems reveals urban mobility patterns and is essential for planning, operation, and optimisation. However, no

Speculative Sampling For Faster Molecular Dynamics

ResearchDGX agent

arXiv:2606.02455v1 Announce Type: new Abstract: Molecular dynamics (MD) is a key tool for simulating the dynamical behavior of atomic systems. However, MD is inherently serial, which makes it difficul

Splatshot: 3D Face Avatar Generation from a Single Unconstrained Photo

ResearchDGX agent

arXiv:2606.01493v1 Announce Type: new Abstract: Reconstructing a photorealistic 3D face avatar from a single unconstrained photograph is challenging: feed-forward 3D Gaussian Splatting (3DGS) models d

Stabilizing Policy Optimization via Logits Convexity

SafetyDGX agent

arXiv:2603.00963v2 Announce Type: replace-cross Abstract: While reinforcement learning (RL) has been central to the recent success of large language models (LLMs), RL optimization is notoriously unsta

Stable Velocity: A Variance Perspective on Flow Matching

Model ReleasesDGX agent

arXiv:2602.05435v2 Announce Type: replace Abstract: While flow matching is elegant, its reliance on single-sample conditional velocities leads to high-variance training targets that destabilize optimi

TAPS: Target-Aware Prefix Tree Selection for Diffusion-Drafted Speculative Decoding

ResearchDGX agent

arXiv:2606.00487v1 Announce Type: new Abstract: Using a diffusion model for parallel drafting is a promising approach for speculative decoding. By predicting tokens at multiple future positions in a s

Tempora: Characterising the Time-Contingent Utility of Online Test-Time Adaptation

ResearchDGX agent

arXiv:2602.06136v2 Announce Type: replace-cross Abstract: Test-time adaptation (TTA) offers a compelling remedy for machine learning (ML) models that degrade under domain shifts, improving generalisat

The Alignment Curse: Modality Alignment Supercharges Audio Attacks via Text Transfer

SafetyDGX agent

arXiv:2602.02557v2 Announce Type: replace-cross Abstract: Recent advances in end-to-end trained omni-models have substantially improved audio capabilities by strengthening text-audio modality alignmen

ThinkSwitch: Context Distillation with LoRA and Weight Interpolation for Specific-Purpose Reasoning Tasks

ResearchDGX agent

arXiv:2606.01080v1 Announce Type: cross Abstract: Large language models often improve on difficult tasks by spending inference-time compute on a reasoning trace before producing the final answer. That

Threshold-Based Exclusive Batching for LLM Inference

HardwareDGX agent

arXiv:2606.00516v1 Announce Type: new Abstract: Mixed batching (MB)--interleaving prefill and decode in a single batch--has become the standard scheduling strategy for large language model (LLM) infer

ToMAP: Training Opponent-Aware LLM Persuaders with Theory of Mind

TutorialsDGX agent

arXiv:2505.22961v3 Announce Type: replace Abstract: Large language models (LLMs) have shown promising potential in persuasion, but existing works on training LLM persuaders are still preliminary. Nota

Towards a General Intelligence and Interface for Wearable Health Data

AgentsDGX agent

arXiv:2605.22759v2 Announce Type: replace Abstract: While ubiquitous wearable sensors capture a wealth of behavioral and physiological information, effectively transforming these signals into personal

Training-Free Composed Video Retrieval via Visual Representation-Guided Video-LLM Reasoning

ResearchDGX agent

arXiv:2606.02321v1 Announce Type: new Abstract: Recent advances in large vision-language models have expanded video retrieval from simple text-based search to more flexible scenarios, where users may

Understanding Identity Continuity in Thermal Video through Scene-Level Consistency

Model ReleasesDGX agent

arXiv:2606.01694v1 Announce Type: cross Abstract: Thermal pedestrian MOT remains challenging because weak appearance cues and frequent detection interruptions cause severe trajectory fragmentation. We

UrbanFusion: Stochastic Multimodal Fusion for Contrastive Learning of Robust Spatial Representations

ResearchDGX agent

arXiv:2510.13774v2 Announce Type: replace-cross Abstract: Forecasting urban phenomena such as housing prices and public health indicators requires the effective integration of various geospatial data.

VideoBrain: Learning Adaptive Frame Sampling for Long Video Understanding

AgentsDGX agent

arXiv:2602.04094v2 Announce Type: replace Abstract: Long-form video understanding remains challenging for Vision-Language Models (VLMs) due to the inherent tension between computational constraints an

When Jokes Cross the Line: Analyzing Regular Humor and Dark Humor in YouTube Shorts

Model ReleasesDGX agent

arXiv:2606.00046v1 Announce Type: cross Abstract: Video platforms such as YouTube have reshaped how users engage with entertainment and information, emphasizing brief, highly engaging content such as

Which Leakage Types Matter? A Quantitative Landscape Across 2,047 Benchmark Datasets

Model ReleasesDGX agent

arXiv:2604.04199v2 Announce Type: replace Abstract: Twenty-eight within-subject counterfactual experiments across 2,047 iid tabular datasets, plus a boundary experiment on 129 temporal datasets, measu

Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025

ResearchDGX agent

arXiv:2606.02255v1 Announce Type: cross Abstract: Human annotation is the empirical foundation of much NLP research, from dataset construction to model evaluation, but papers often leave unclear who p

← Previous
1…542543544545546…1061
Next →