AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
2 Jun 2026

Interpretable Self-Supervised Learning via Representer Landmarks and Nystrom Approximation

SafetyDGX agent

arXiv:2509.24467v3 Announce Type: replace Abstract: Self-supervised learning (SSL) learns representations from massive unlabeled data, yet the resulting models typically operate as black boxes, necess

Interpreto: An Explainability Library for Transformers

ResearchDGX agent

arXiv:2512.09730v3 Announce Type: replace Abstract: Interpreto is an open-source Python library for interpreting HuggingFace language models, from early BERT variants to LLMs. It provides two compleme

Is Zero-Shot Super-Resolution Possible in Operator Learning?

ResearchDGX agent

arXiv:2606.00296v1 Announce Type: cross Abstract: Neural operators are often reported to exhibit zero-shot super-resolution, a phenomenon in which a model trained on coarse grids produces accurate pre

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

It is difficult to know how good MAI-Thinking-1 is from the scores alone (like weirdly low GPQA & Terminal Bench 2.0) But Microsoft makes it…

ApplicationsDGX agent

It is difficult to know how good MAI-Thinking-1 is from the scores alone (like weirdly low GPQA & Terminal Bench 2.0) But Microsoft makes it really hard to try its models upon release (a general issue

Iteris: Agentic Research Loops for Computational Mathematics

AgentsDGX agent

arXiv:2606.02484v1 Announce Type: new Abstract: Recent advances in large language models and agentic AI systems have enabled significant progress in mathematical discovery, from solving competition pr

Join the livestream to hear from our team members @karan4d and @yoniebans live from @nvidia! https://www.youtube.com/watch?v=pgQDbRMa2Eg

HardwareDGX agent

Nous Research is hosting a livestream featuring team members Karan and Yoni speaking from NVIDIA, likely discussing AI research, model development, or collaboration between Nous Research and NVIDIA. T

KACE: Knowledge-Adaptive Context Engineering for Mathematical Reasoning

ResearchDGX agent

arXiv:2606.00532v1 Announce Type: new Abstract: Context engineering can improve large language models without updating their weights, but mathematical reasoning exposes a key limitation: feedback accu

Last Layer Logits to Logic: Empowering LLMs with Logic-Consistent Structured Knowledge Reasoning

TutorialsDGX agent

arXiv:2511.07910v2 Announce Type: replace Abstract: Large Language Models (LLMs) achieve excellent performance in natural language reasoning tasks through pre-training on vast unstructured text, enabl

LeARN: Learnable and Adaptive Representations for Nonlinear Dynamics in System Identification

AgentsDGX agent

arXiv:2412.12036v2 Announce Type: replace Abstract: System identification, the process of deriving mathematical models of dynamical systems from observed input-output data, has undergone a paradigm sh

Learning Hamiltonian Dynamics at Scale: A Differential-Geometric Approach

ResearchDGX agent

arXiv:2509.24627v2 Announce Type: replace Abstract: Embedding physical intuition into network architectures allows the learning of dynamics that enforce fundamental properties, such as energy conserva

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.02132v1 Announce Type: new Abstract: Agentic reinforcement learning can induce tool abuse, where models overuse external tools even for queries solvable by internal reasoning. Existing appr

LLM Trainer: Automated Robotic Data Generation via Demonstration Augmentation using LLMs

SafetyDGX agent

arXiv:2509.20070v2 Announce Type: replace Abstract: We present LLM Trainer, a fully automated pipeline that leverages the world knowledge of Large Language Models (LLMs) to transform a small number of

LLMs Need Encoders for Semantic IDs Too

TutorialsDGX agent

arXiv:2606.00324v1 Announce Type: cross Abstract: Multimodal LLMs use dedicated encoders to bridge non-language modalities (vision encoders for images, depth models for audio codec tokens) because raw

Lodestar: An Online-Learning LLM Inference Router

HardwareDGX agent

arXiv:2606.00946v1 Announce Type: cross Abstract: Efficiently serving large language model (LLM) inference tasks is crucial both for user-perceived latency such as time-to-first-token (TTFT) and for G

Machine Learning for Coding Retail Product Names to Consumer-Price Categories: A Rule-plus-Bag-of-Words Pipeline with Reliability-Weighted Human-in-the-Loop Labeling

ResearchDGX agent

arXiv:2606.02004v1 Announce Type: new Abstract: Consumer-price measurement increasingly draws on alternative data sources -- scanner, web-scraped, and transaction/receipt data. A recurring obstacle is

Measurement Geometry and Design for Trustworthy Generative Inverse Problems

SafetyDGX agent

arXiv:2606.02309v1 Announce Type: cross Abstract: Generative models are increasingly used as priors for inverse problems, but their ability to produce realistic images creates a basic trust problem: a

MESA: Improving MoE Safety Alignment via Decentralized Expertise

SafetyDGX agent

arXiv:2606.00651v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) architectures scale Large Language Models (LLMs) efficiently, enabling greater capacity with reduced computational cost by dy

Minimax-Optimal Policy Regret in Partially Observable Markov Games

SafetyDGX agent

arXiv:2606.02363v1 Announce Type: new Abstract: We study sequential decision-making in partially observable environments against strategic, adaptive opponents, modeled as partially observable Markov g

MINTS: Minimalist Thompson Sampling

ResearchDGX agent

arXiv:2606.01655v1 Announce Type: cross Abstract: The Bayesian paradigm offers principled tools for sequential decision-making under uncertainty, but its reliance on a probabilistic model for all para

Normality-Preserving Continual Industrial Anomaly Detection via Orthogonal LoRA Banks

ResearchDGX agent

arXiv:2606.02042v1 Announce Type: new Abstract: Continual industrial anomaly detection with diffusion models suffers from historical normality prior drift and catastrophic forgetting. Existing continu

Not All Flips Are Conformity: Decomposing Stance Convergence in Multi-Agent LLM Debate

AgentsDGX agent

arXiv:2606.00820v1 Announce Type: new Abstract: Multi-agent debate (MAD) is a promising strategy for improving LLM reasoning, but when agents converge on a shared answer, it is unclear whether that co

Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis

TutorialsDGX agent

arXiv:2606.02510v1 Announce Type: new Abstract: Constructing faithful 4D worlds from LiDAR-acquired sequences is crucial for embodied AI, yet current generative frameworks apply uniform modeling capac

Not What, But How: A Communicative Audit of LLM Response Framing

ResearchDGX agent

arXiv:2606.02493v1 Announce Type: new Abstract: Large language models (LLMs) are being increasingly used to answer subjective, information-seeking questions, where users are sensitive to how responses

Ollama can't list this C# game script, because it might cause destruction.

Local AiDGX agent

A Reddit post discussing an issue where Ollama (an AI model tool) refuses to process or list a C# game script due to safety concerns about potential destructive code. The post likely explores the limi

On Effectiveness and Efficiency of Agentic Tool-calling and RL Training

SafetyDGX agent

arXiv:2606.00135v1 Announce Type: cross Abstract: Tool-calling is a central component of modern large language model (LLM) agents, equipping them with skills beyond their parametric knowledge. This pa

On the Collapse of Generative Paths: A Criterion and Correction for Diffusion Steering

ResearchDGX agent

arXiv:2512.10339v2 Announce Type: replace Abstract: Inference-time steering adapts pretrained diffusion and flow models to new tasks without retraining, often utilizing ratio-of-densities construction

One Channel to Rule Them All: Rethinking Input Representation for Visual Place Recognition

Local AiDGX agent

arXiv:2606.00936v1 Announce Type: new Abstract: Visual Place Recognition (VPR) is fundamental to long-term robot localization and SLAM, yet current systems overwhelmingly rely on RGB input, implicitly

OP-LoRA: The Blessing of Dimensionality

ResearchDGX agent

arXiv:2412.10362v2 Announce Type: replace-cross Abstract: Low-rank adapters (LoRA) enable finetuning of large models with only a small number of parameters. However, they often suffer from an ill-cond

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence

AgentsDGX agent

arXiv:2603.14771v3 Announce Type: replace Abstract: Large Language Model (LLM)-based Collective Intelligence (CI) presents a promising approach to overcoming the data wall and continuously boosting th

PaCo-VLA: Passivity-Shielded Compliance Prior for Contact-Rich Vision-Language-Action Manipulation

ApplicationsDGX agent

arXiv:2606.00515v1 Announce Type: cross Abstract: Contact-rich manipulation demands both high-level semantic reasoning and the safe regulation of high-frequency contact dynamics. While Vision-Language

PaCX-MAE: Physiology-Augmented Chest X-Ray Masked Autoencoder

ResearchDGX agent

arXiv:2606.01537v1 Announce Type: new Abstract: Clinical diagnosis often requires combining imaging with physiological measurements, yet deployed models typically operate on unimodal data. We present

Paradoxical noise preference in RNNs

SafetyDGX agent

arXiv:2601.04539v2 Announce Type: replace-cross Abstract: In recurrent neural networks (RNNs) used to model biological neural networks, noise is typically introduced during training to emulate biologi

Parametric Social Identity Injection and Diversification in Public Opinion Simulation

ApplicationsDGX agent

arXiv:2603.16142v2 Announce Type: replace Abstract: Large language models (LLMs) have recently been adopted as synthetic agents for public opinion simulation, offering a promising alternative to costl

Partial Fairness Awareness: Belief-Guided Strategic Mechanism for Strategic Agents

SafetyDGX agent

arXiv:2606.00826v1 Announce Type: new Abstract: Strategic machine learning investigates scenarios where agents manipulate their features to receive favorable decisions from predictive models. To addre

Physics-Aware Linearized ADMM and Its Unrolling

ResearchDGX agent

arXiv:2606.01652v1 Announce Type: cross Abstract: Recently, partial differential equations (PDEs) have been used to directly model the measurement process in signal processing, although their evaluati

PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation

AgentsDGX agent

arXiv:2602.01662v4 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have enabled increasing progress in real-world robot manipulation. However, long-horizon manipulati

PMC-InterCPT: Rethinking Biomedical Interleaved Data for Multimodal Continued Pretraining

ResearchDGX agent

arXiv:2606.01049v1 Announce Type: new Abstract: Large-scale biomedical image-text datasets extracted from scientific literature provide valuable resources for medical multimodal model training. These

Predicting Future Utility: Global Combinatorial Optimization for Task-Agnostic KV Cache Eviction

HardwareDGX agent

arXiv:2602.08585v2 Announce Type: replace-cross Abstract: Given the quadratic complexity of attention, KV cache eviction is vital to accelerate model inference. Current KV cache eviction methods typic

Principle-Evolvable Scientific Discovery via Uncertainty Minimization

ResearchDGX agent

arXiv:2602.06448v2 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based scientific agents have accelerated scientific discovery, yet they often suffer from significant inefficiencie

ProtoAda: Prototype-Guided Adaptive Adapter Expansion and Geometric Consolidation for Multimodal Continual Instruction Tuning

ApplicationsDGX agent

arXiv:2606.02576v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to continually a

PSG-Nav: Probabilistic Scene Graph Navigation via Multiverse Decision Making

ResearchDGX agent

arXiv:2606.01313v1 Announce Type: cross Abstract: Open-vocabulary navigation requires embodied agents to manage significant perception uncertainty stemming from semantic ambiguity and model errors. Ho

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography

AgentsDGX agent

arXiv:2604.15231v2 Announce Type: replace Abstract: Vision-language models (VLM) have markedly advanced AI-driven interpretation and reporting of complex medical imaging, such as computed tomography (

Rank-Constrained Deep Matrix Completion for Group Recommendation

ApplicationsDGX agent

arXiv:2606.01948v1 Announce Type: cross Abstract: The growing popularity of group activities has increased the need for methods that provide recommendations to groups of users given their individual p

Reason, Retrieve, Re-rank: A Zero-Shot Reasoning-Aware Framework for Composed Video Retrieval

ResearchDGX agent

arXiv:2606.00910v1 Announce Type: new Abstract: Composed Video Retrieval (CoVR) seeks the target video that results from applying a free-form textual modification to a reference video. We address the

Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief

SafetyDGX agent

arXiv:2606.00680v1 Announce Type: new Abstract: Offline reinforcement learning (RL) aims to optimize policies from pre-collected datasets. A bottleneck of this paradigm is managing epistemic uncertain

Repurposing Adversarial Perturbations for Continual Learning: From Defense to Active Alignment

SafetyDGX agent

arXiv:2606.02322v1 Announce Type: cross Abstract: In dynamic environments, large language models need to keep adapting to new tasks, but continual learning often suffers from forgetting, limited trans

Retrieve What's Missing: Coverage-Maximizing Retrieval for Consistent Long Video Generation

ResearchDGX agent

arXiv:2606.02479v1 Announce Type: new Abstract: Maintaining long-term geometric consistency remains challenging for long-horizon autoregressive video generation. Memory-augmented generative models add

Robust Integrated Planning and Control for Quadrotors in Dynamic Environments via NMPC with CBF Penalties

SafetyDGX agent

arXiv:2606.01038v1 Announce Type: new Abstract: This paper presents a new robust integrated planning and control (IPC) strategy for multirotor uncrewed aerial vehicles. We propose a nonlinear model pr

Safe-Subspace Pseudo-Label Refinement for Source-Free Graph Domain Adaptation

ApplicationsDGX agent

arXiv:2606.00808v1 Announce Type: new Abstract: Source-free graph domain adaptation (SF-GDA) aims to adapt source-trained graph models to unlabeled target graphs when source graphs are no longer acces

Scaling Agentic Capabilities via Grounded Interaction Synthesis

AgentsDGX agent

arXiv:2606.02001v1 Announce Type: new Abstract: General agentic intelligence hinges on the ability to interact with diverse real-world tools to complete complex tasks, a capability fundamentally tied

Self-Improving Small Object Grounding in LVLMs

ResearchDGX agent

arXiv:2606.01612v1 Announce Type: new Abstract: Can internal attention patterns in Large Vision Language Models (LVLMs) identify reliable small-object boxes without fine-tuning? In this work, we provi

Situation-Aware Interactive MPC Switching for Autonomous Driving

AgentsDGX agent

arXiv:2512.06182v2 Announce Type: replace Abstract: Autonomous driving in interactive traffic scenarios remains challenging because of the mutual influence among vehicles and the inherent uncertainty

Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills

HardwareDGX agent

arXiv:2503.05641v4 Announce Type: replace-cross Abstract: Combining existing pre-trained LLMs is a promising approach for diverse reasoning tasks. However, task-level expert selection is often too coa

'Skill issues'': data-centric optimization of lakehouse agents

AgentsDGX agent

arXiv:2606.01185v1 Announce Type: new Abstract: Coding agents are becoming users of data infrastructure, but their success depends not only on model quality: it also depends on the skills and environm

SN-WER: Script-Normalized WER for Multi-Script Indic ASR Evaluation

ResearchDGX agent

arXiv:2606.02548v1 Announce Type: new Abstract: Word Error Rate (WER) is the dominant metric for automatic speech recognition (ASR), but it can overestimate errors when references and hypotheses encod

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that th…

SafetyDGX agent

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that the environment can maintain reliably, it probably doesn't bel

Statistical Guarantees for Reasoning Probes on Looped Boolean Circuits

ResearchDGX agent

arXiv:2602.03970v3 Announce Type: replace-cross Abstract: We study the statistical behavior of reasoning probes in a stylized model of iterative computation inspired by neural algorithmic reasoning. T

SWARD: Stochastic Window-Attention-Based Relational Distillation for Cross-Architectural Semantic Segmentation

SafetyDGX agent

arXiv:2606.00999v1 Announce Type: new Abstract: Large-scale vision foundation models have driven substantial gains on dense prediction tasks such as semantic segmentation, but their size makes deploym

T-POP: Test-Time Personalization with Online Preference Feedback

SafetyDGX agent

arXiv:2509.24696v2 Announce Type: replace-cross Abstract: Personalizing large language models (LLMs) to individual user preferences is a critical step beyond generating generically helpful responses.

Tackling the Root of Misinformation by Teaching Laypeople about Logical Fallacies via Socratic Questioning and Critical Argumentation

TutorialsDGX agent

arXiv:2606.01020v1 Announce Type: new Abstract: Identifying logical fallacies in everyday discourse is challenging for many people. This challenge is amplified in the era of Large Language Models (LLM

← Previous
1…757758759760761…1018
Next →