AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
12 May 2026

Hyperspherical Autoencoder for High-Fidelity Image Reconstruction and Generation

SafetyDGX agent

arXiv:2601.22904v2 Announce Type: replace-cross Abstract: Recent studies have explored using pretrained Vision Foundation Models (VFMs) such as DINO for generative autoencoders, showing strong generat

Interpretable Machine Learning for Football Performance Analysis: Evidence of Limited Transferability from Elite Leagues to University Competition

ResearchDGX agent

arXiv:2605.10796v1 Announce Type: new Abstract: Machine learning has become increasingly prevalent in football performance analysis, yet most studies prioritize predictive accuracy while implicitly as

Key Coverage Matters: Semi-Structured Extraction of OCR Clinical Reports

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.09440v1 Announce Type: cross Abstract: Clinical reports are often fragmented across healthcare institutions because privacy regulations and data silos limit direct information sharing. When

Language-Conditioned Visual Grounding with CLIP Multilingual

ResearchDGX agent

arXiv:2605.09060v1 Announce Type: new Abstract: Multilingual vision-language models exhibit systematic performance gaps across languages, but the mechanism remains ambiguous: cross-language divergence

Learning More from Less: Exploiting Counterfactuals for Data-Efficient Chart Understanding

ResearchDGX agent

arXiv:2605.10855v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in chart understanding, largely driven by supervised fine-tuning (SFT) on increasing

Learning Rate Scheduling with Matrix Factorization for Private Training

ResearchDGX agent

arXiv:2511.17994v2 Announce Type: replace Abstract: We study differentially private model training with stochastic gradient descent under learning rate scheduling and correlated noise. Although correl

Learning When to Jump for Off-road Navigation

SafetyDGX agent

arXiv:2602.00877v2 Announce Type: replace Abstract: Low speed does not always guarantee safety in off-road driving. For instance, crossing a ditch may be risky at a low speed due to the risk of gettin

Lecture Notes on Statistical Physics and Neural Networks

ResearchDGX agent

arXiv:2605.06394v1 Announce Type: cross Abstract: These lecture notes introduce some topics of classical statistical physics, particularly those that are relevant for neural networks and deep learning

Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes

HardwareDGX agent

arXiv:2505.02184v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for generating parallel scientific codes, with a primary focus on generating functionally correct

LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation

ApplicationsDGX agent

arXiv:2605.10593v1 Announce Type: new Abstract: We demonstrate LLARS (LLM Assisted Research System), an open-source platform that bridges the gap between domain experts and developers for building LLM

LLM Advertisement based on Neuron Auctions

SafetyDGX agent

arXiv:2605.08326v1 Announce Type: cross Abstract: As Large Language Models (LLMs) transition into conversational agents, generative advertising emerges as a crucial monetization strategy. However, emb

LLM-guided Semi-Supervised Approaches for Social Media Crisis Data Classification

ApplicationsDGX agent

arXiv:2605.08448v1 Announce Type: new Abstract: Semi-supervised learning approaches have been investigated as a means to enhance the analysis of social media data in disaster management contexts. In t

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

HardwareDGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

Manifold scores 7,700 MCP servers in Manifest expansion aimed at agent security teams

AgentsDGX agent

Artificial intelligence detection and response platform startup Manifold Security Inc. today announced an expansion of its Manifest supply chain intelligence tool to cover Model Context Protocol serve

Masked Generative Transformer Is What You Need for Image Editing

ResearchDGX agent

arXiv:2605.10859v1 Announce Type: new Abstract: Diffusion models dominate image editing, yet their global denoising mechanism entangles edited regions with surrounding context, causing modifications t

Memorize Theorems, Not Instances: Probing SFT Generalization through Mathematical Reasoning

ResearchDGX agent

arXiv:2605.09270v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) is widely used for task-specific adaptation, yet recent work shows it systematically undermines reasoning generalization.

MeshFIM: Local Low-Poly Mesh Editing via Fill-in-the-Middle Autoregressive Generation

ResearchDGX agent

arXiv:2605.08744v1 Announce Type: cross Abstract: Autoregressive (AR) models can generate high-quality low-poly meshes from point clouds, but they still operate in an all-or-nothing manner: when a loc

MIND-Skill: Quality-Guaranteed Skill Generation via Multi-Agent Induction and Deduction

AgentsDGX agent

arXiv:2605.08670v1 Announce Type: new Abstract: Large language model (LLM) powered AI agents have emerged as a promising paradigm for autonomous problem-solving, yet they continue to struggle with com

Mind the Gap No More: Achieving Zero-Gap Multimodal Integration via One Tokenizer

ResearchDGX agent

arXiv:2602.12286v2 Announce Type: replace-cross Abstract: A central challenge in developing Multimodal Large Language Models (MLLMs) is effectively integrating heterogeneous inputs into a cohesive rea

Mixture of Experts for Recognizing Depression from Interview and Reading Tasks

ResearchDGX agent

arXiv:2502.20213v2 Announce Type: replace Abstract: Depression is a mental disorder and can cause a variety of symptoms, including psychological, physical, and social. Speech has been proved an object

Multi-Armed Bandits With Best-Action Queries

ResearchDGX agent

arXiv:2605.08287v1 Announce Type: cross Abstract: We study multi-armed bandits (MABs) augmented with best-action queries, in which the learner may additionally query an oracle that reveals the best ar

Muon Does Not Converge on Convex Lipschitz Functions

ResearchDGX agent

arXiv:2605.08980v1 Announce Type: new Abstract: Muon and its variants have shown strong empirical performance in a variety of deep learning tasks. Existing convergence analyses of Muon rely on smoothn

Nano-U: Efficient Terrain Segmentation for Tiny Robot Navigation

AgentsDGX agent

arXiv:2605.10210v1 Announce Type: cross Abstract: Terrain segmentation is a fundamental capability for autonomous mobile robots operating in unstructured outdoor environments. However, state-of-the-ar

NCO: A Versatile Plug-in for Handling Negative Constraints in Decoding

ResearchDGX agent

arXiv:2605.10065v1 Announce Type: cross Abstract: Controlling Large Language Models (LLMs) to prevent the generation of undesirable content, such as profanity and personally identifiable information (

Nectar: Neural Estimation of Cached-Token Attention via Regression

ApplicationsDGX agent

arXiv:2605.09778v1 Announce Type: cross Abstract: Evaluating softmax attention over a fixed long context requires reading every cached key-value pair for each new query token. For a given context (a b

NEXT: Multi-Grained Mixture of Experts via Text-Modulation for Multi-Modal Object Re-Identification

ApplicationsDGX agent

arXiv:2505.20001v5 Announce Type: replace Abstract: Multi-modal object Re-IDentification (ReID) aims to obtain complete identity features across heterogeneous modalities. However, most existing method

NEXUS: Continual Learning of Symbolic Constraints for Safe and Robust Embodied Planning

SafetyDGX agent

arXiv:2605.09387v1 Announce Type: new Abstract: While Large Language Models (LLMs) have catalyzed progress in embodied intelligence, a fundamental gap between their inherent probabilistic uncertainty

Non-intrusive Body Composition Assessment from Full-body mmWave Scans

ResearchDGX agent

arXiv:2605.08306v1 Announce Type: cross Abstract: Body composition assessment (BCA) provides detailed information about the distribution of different tissue types in the body, enabling more precise ch

OsteoFlow: Lyapunov-Guided Flow Distillation for Predicting Bone Remodeling after Mandibular Reconstruction

ResearchDGX agent

arXiv:2603.22421v2 Announce Type: replace Abstract: Predicting long-term bone remodeling after mandibular reconstruction would be of great clinical benefit, yet standard generative models struggle to

Outlier-robust Diffusion Posterior Sampling for Bayesian Inverse Problems

TutorialsDGX agent

arXiv:2602.02045v2 Announce Type: replace Abstract: Diffusion models have emerged as powerful learned priors for Bayesian inverse problems (BIPs). Diffusion-based solvers rely on a presumed likelihood

Outlier-Robust Diffusion Solvers for Inverse Problems

ApplicationsDGX agent

arXiv:2605.09477v1 Announce Type: cross Abstract: Methods based on diffusion models (DMs) for solving inverse problems (IPs) have recently achieved remarkable performance. However, DM-based methods ty

PAAC: Privacy-Aware Agentic Device-Cloud Collaboration

Local AiDGX agent

arXiv:2605.08646v1 Announce Type: cross Abstract: Large language model (LLM) agents face a structural tension: cloud agents provide strong reasoning but expose user data, while on-device agents preser

PaceVGGT: Pre-Alternating-Attention Token Pruning for Visual Geometry Transformers

ResearchDGX agent

arXiv:2605.08371v1 Announce Type: new Abstract: Visual Geometry Transformer (VGGT) is a strong feed-forward model for multiple 3D tasks, but its Alternating-Attention (AA) stack scales quadratically i

Path-Coupled Bellman Flows for Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2605.08253v1 Announce Type: cross Abstract: Distributional reinforcement learning (DRL) models the full return distribution, but existing finite-support or quantile-based methods rely on project

PathISE: Learning Informative Path Supervision for Knowledge Graph Question Answering

ResearchDGX agent

arXiv:2605.10791v1 Announce Type: new Abstract: Knowledge Graph Question Answering (KGQA) aims to answer user questions by reasoning over Knowledge Graphs (KGs). Recent KGQA methods mainly follow the

Performance and Energy Trade-Off Analysis of Hierarchical Federated Learning for Plant Disease Classification

ResearchDGX agent

arXiv:2605.08121v1 Announce Type: cross Abstract: Early detection of plant diseases is critical for improving crop productivity, while it also facilitates the foundations of precision agriculture. Rec

Pixal3D: Pixel-Aligned 3D Generation from Images

ResearchDGX agent

arXiv:2605.10922v1 Announce Type: new Abstract: Recent advances in 3D generative models have rapidly improved image-to-3D synthesis quality, enabling higher-resolution geometry and more realistic appe

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation

SafetyDGX agent

arXiv:2605.10118v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated exceptional general reasoning capabilities. However, their performance in embodied navigation remains hi

Practical Scaling Laws: Converting Compute into Performance in a Data-Constrained World

ResearchDGX agent

arXiv:2605.09189v1 Announce Type: new Abstract: The scaling laws guiding modern model training were calibrated for a single regime: data-rich, single-epoch pretraining. The dominant such scaling law f

ProDG: Prototypes for Data-Free Generative Post-Hoc Explainability

ResearchDGX agent

arXiv:2605.08858v1 Announce Type: new Abstract: Ante-hoc interpretability methods based on prototypes provide highly accurate explanations by utilizing the intuitive 'this looks like that' reasoning p

PruneTIR: Inference-Time Tool Call Pruning for Effective yet Efficient Tool-Integrated Reasoning

TutorialsDGX agent

arXiv:2605.09931v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) enables large language models (LLMs) to enhance their capabilities by interacting with external tools, such as code in

PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents

Local AiDGX agent

arXiv:2605.08468v1 Announce Type: cross Abstract: Local LLM-based coding agents increasingly work in settings where correctness is earned through execution feedback, persistent state, and bounded repa

RDKV: Rate-Distortion Bit Allocation for Joint Eviction and Quantization of the KV Cache

ResearchDGX agent

arXiv:2605.08317v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong performance across diverse tasks, but their inference with long input contexts is bottlenecked by memor

Re-Triggering Safeguards within LLMs for Jailbreak Detection

SafetyDGX agent

arXiv:2605.10611v1 Announce Type: cross Abstract: This paper proposes a jailbreaking prompt detection method for large language models (LLMs) to defend against jailbreak attacks. Although recent LLMs

Reasoning-Aware Training for Time Series Forecasting

ResearchDGX agent

arXiv:2605.08625v1 Announce Type: cross Abstract: Time Series Foundation Models (TSFMs) excel at numerical forecasting but operate as black boxes lacking qualitative reasoning. Conversely, applying LL

Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge

SafetyDGX agent

arXiv:2605.10805v1 Announce Type: new Abstract: Reasoning-capable large language models (LLMs) have recently been adopted as automated judges, but their benefits and costs in LLM-as-a-Judge settings r

Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning

SafetyDGX agent

arXiv:2605.09614v1 Announce Type: new Abstract: Long chain-of-thought (CoT) reasoning improves large vision--language models, but visual information often fades during generation, limiting long-horizo

Reinforcing Multimodal Reasoning Against Visual Degradation

SafetyDGX agent

arXiv:2605.09262v1 Announce Type: cross Abstract: Reinforcement Learning has significantly advanced the reasoning capabilities of Multimodal Large Language Models (MLLMs), yet the resulting policies r

Remember to Forget: Gated Adaptive Positional Encoding

SafetyDGX agent

arXiv:2605.10414v1 Announce Type: new Abstract: Rotary Positional Encoding (RoPE) is widely used in modern large language models. However, when sequences are extended beyond the range seen during trai

RL Fine-Tuning Heals OOD Forgetting in SFT

ResearchDGX agent

arXiv:2509.12235v3 Announce Type: replace-cross Abstract: Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) is a standard post-training recipe for improving Large Language Models (L

Robust Multi-Agent LLMs under Byzantine Faults

AgentsDGX agent

arXiv:2605.09076v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly collaborate over peer-to-peer networks to improve their reliability. However, these same interactions c

RUBEN: Rule-Based Explanations for Retrieval-Augmented LLM Systems

SafetyDGX agent

arXiv:2605.10862v1 Announce Type: new Abstract: This paper demonstrates RUBEN, an interactive tool for discovering minimal rules to explain the outputs of retrieval-augmented large language models (LL

Sanity Checks for Long-Form Hallucination Detection

ResearchDGX agent

arXiv:2605.08346v1 Announce Type: cross Abstract: Hallucination detection methods for large language models increasingly operate on chain-of-thought reasoning traces, yet it remains unclear whether th

Scalable Mamba-Based Message-Passing Neural Decoder for Error-Correcting Codes

ResearchDGX agent

arXiv:2605.10681v1 Announce Type: cross Abstract: Forward error correction is essential for reliable communication over noisy channels. Attention-based model-free neural decoders have shown strong per

SDG-MoE: Signed Debate Graph Mixture-of-Experts

ResearchDGX agent

arXiv:2605.08322v1 Announce Type: cross Abstract: Sparse MoE models achieve a good balance between capacity and compute by routing each token to a small subset of experts. However, in most MoE archite

SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization

AgentsDGX agent

arXiv:2605.08382v1 Announce Type: cross Abstract: LLM coding agents now generate code at an unprecedented scale, yet LLM-generated code introduces cybersecurity vulnerabilities into codebases without

Shields to Guarantee Probabilistic Safety in MDPs

SafetyDGX agent

arXiv:2605.10888v1 Announce Type: cross Abstract: Shielding is a prominent model-based technique to ensure safety of autonomous agents. Classical shielding aims to ensure that nothing bad ever happens

Signal from Structure: Exploiting Submodular Upper Bounds in Generative Flow Networks

TutorialsDGX agent

arXiv:2601.21061v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets; GFNs) are a class of generative models that learn to sample compositional objects proportionally to their a pri

Simulating Complex Multi-Turn Tool Calling Interactions in Stateless Execution Environments

AgentsDGX agent

arXiv:2601.19914v2 Announce Type: replace-cross Abstract: Synthetic data has proven itself to be a valuable resource for tuning smaller, cost-effective language models to handle the complexities of mu

SkillMAS: Skill Co-Evolution with LLM-based Multi-Agent System

AgentsDGX agent

arXiv:2605.09341v1 Announce Type: cross Abstract: Large language model (LLM) agent systems are increasingly expected to improve after deployment, but existing work often decouples two adaptation targe

← Previous
1…778779780781782…1017
Next →