AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
14 Apr 2026

Muon^2: Boosting Muon via Adaptive Second-Moment Preconditioning

Model ReleasesDGX agent

arXiv:2604.09967v1 Announce Type: cross Abstract: Muon has emerged as a promising optimizer for large-scale foundation model pre-training by exploiting the matrix structure of neural network updates t

Nano-EmoX: Unifying Multimodal Emotional Intelligence from Perception to Empathy

ResearchDGX agent

arXiv:2603.02123v3 Announce Type: replace Abstract: The development of affective multimodal language models (MLMs) has long been constrained by a gap between low-level perception and high-level intera

Natural Gradient Gaussian Approximation Filter on Lie Groups for Robot State Estimation

Model ReleasesDGX agent

arXiv:2604.10057v1 Announce Type: new Abstract: Accurate state estimation for robotic systems evolving on Lie group manifolds, such as legged robots, is a prerequisite for achieving agile control. How

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

NetworkNet: A Deep Neural Network Approach for Random Networks with Sparse Nodal Attributes and Complex Nodal Heterogeneity

ResearchDGX agent

arXiv:2604.11673v1 Announce Type: cross Abstract: Heterogeneous network data with rich nodal information become increasingly prevalent across multidisciplinary research, yet accurately modeling comple

NVIDIA Ising Introduces AI-Powered Workflows to Build Fault-Tolerant Quantum Systems

HardwareDGX agent

NVIDIA Ising is the world's first family of open-source quantum AI models, designed to help researchers and enterprises build quantum processors capable of running useful applications. The family span

ODUTQA-MDC: A Task for Open-Domain Underspecified Tabular QA with Multi-turn Dialogue-based Clarification

Model ReleasesDGX agent

arXiv:2604.10159v1 Announce Type: new Abstract: The advancement of large language models (LLMs) has enhanced tabular question answering (Tabular QA), yet they struggle with open-domain queries exhibit

Online Learning-Enhanced High Order Adaptive Safety Control

SafetyDGX agent

arXiv:2511.19651v2 Announce Type: replace Abstract: Control barrier functions (CBFs) are an effective model-based tool to formally certify the safety of a system. With the growing complexity of modern

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents dep…

Model ReleasesDGX agent

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents deploy has easy configs to let users customize their harness and

Pacing Opinion Polarization via Graph Reinforcement Learning

ApplicationsDGX agent

arXiv:2602.23390v2 Announce Type: replace-cross Abstract: Opinion polarization moderation has been studied mainly as an analytical optimization problem under the Friedkin Johnson FJ model, where inter

PanoSAMic: Panoramic Image Segmentation from SAM Feature Encoding and Dual View Fusion

ResearchDGX agent

arXiv:2601.07447v2 Announce Type: replace Abstract: Existing image foundation models are not optimized for spherical images having been trained primarily on perspective images. PanoSAMic integrates th

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

Model ReleasesDGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

PAS: Estimating the target accuracy before domain adaptation

ResearchDGX agent

arXiv:2604.09863v1 Announce Type: cross Abstract: The goal of domain adaptation is to make predictions for unlabeled samples from a target domain with the help of labeled samples from a different but

Poisoning with A Pill: Circumventing Detection in Federated Learning

ResearchDGX agent

arXiv:2407.15389v2 Announce Type: replace Abstract: Without direct access to the client's data, federated learning (FL) is well-known for its unique strength in data privacy protection among existing

Post-Processing Methods for Improving Accuracy in MRI Inpainting

Local AiDGX agent

arXiv:2510.15282v2 Announce Type: replace-cross Abstract: Magnetic Resonance Imaging (MRI) is the primary imaging modality used in the diagnosis, assessment, and treatment planning for brain pathologi

Precision Synthesis of Multi-Tracer PET via VLM-Modulated Rectified Flow for Stratifying Mild Cognitive Impairment

Model ReleasesDGX agent

arXiv:2604.11176v1 Announce Type: new Abstract: The biological definition of Alzheimer's disease (AD) relies on multi-modal neuroimaging, yet the clinical utility of positron emission tomography (PET)

PRIX: Learning to Plan from Raw Pixels for End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2507.17596v3 Announce Type: replace-cross Abstract: While end-to-end autonomous driving models show promising results, their practical deployment is often hindered by large model sizes, a relian

Probabilistic Prediction of Neural Dynamics via Autoregressive Flow Matching

TutorialsDGX agent

arXiv:2604.11178v1 Announce Type: cross Abstract: Forecasting neural activity in response to naturalistic stimuli remains a key challenge for understanding brain dynamics and enabling downstream neuro

Psychological Concept Neurons: Can Neural Control Bias Probing and Shift Generation in LLMs?

Local AiDGX agent

arXiv:2604.11802v1 Announce Type: new Abstract: Using psychological constructs such as the Big Five, large language models (LLMs) can imitate specific personality profiles and predict a user's persona

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

Model ReleasesDGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought

SafetyDGX agent

arXiv:2506.16796v4 Announce Type: replace Abstract: Real-World Image Super-Resolution is one of the most challenging task in image restoration. However, existing methods struggle with an accurate unde

Reasoning as Gradient: Scaling MLE Agents Beyond Tree Search

Model ReleasesDGX agent

arXiv:2603.01692v3 Announce Type: replace-cross Abstract: LLM-based agents for machine learning engineering (MLE) predominantly rely on tree search, a form of gradient-free optimization that uses scal

ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization

Model ReleasesDGX agent

arXiv:2604.10520v1 Announce Type: cross Abstract: As Large Language Models (LLMs) have become capable of generating long and descriptive code summaries, accurate and reliable evaluation of factual con

Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

Model ReleasesDGX agent

arXiv:2604.11554v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language mode

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

Model ReleasesDGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

RL makes MLLMs see better than SFT

Model ReleasesDGX agent

arXiv:2510.16333v2 Announce Type: replace Abstract: A dominant assumption in Multimodal Language Model (MLLM) research is that its performance is largely inherited from the LLM backbone, given its imm

Robust Fair Disease Diagnosis in CT Images

Model ReleasesDGX agent

arXiv:2604.09710v1 Announce Type: new Abstract: Automated diagnosis from chest CT has improved considerably with deep learning, but models trained on skewed datasets tend to perform unevenly across pa

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

Local AiDGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

Model ReleasesDGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

Model ReleasesDGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding

Model ReleasesDGX agent

arXiv:2604.11122v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated immense potential in Earth observation. However, the massive visual tokens generated when p

Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence

SafetyDGX agent

arXiv:2604.10404v1 Announce Type: cross Abstract: Edge-based multimodal medical monitoring requires models that balance diagnostic accuracy with severe energy constraints. Continuous acquisition of EC

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

Model ReleasesDGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

Steered LLM Activations are Non-Surjective

SafetyDGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

Model ReleasesDGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

ResearchDGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

Teaching Robots to Interpret Social Interactions through Lexically-guided Dynamic Graph Learning

ResearchDGX agent

arXiv:2604.10895v1 Announce Type: cross Abstract: For a robot to be called socially intelligent, it must be able to infer users internal states from their current behaviour, predict the users future b

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

Model ReleasesDGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2510.15552v3 Announce Type: replace-cross Abstract: Large language models (LLMs) still struggle with multi-hop reasoning over knowledge-graphs (KGs), and we identify a previously overlooked stru

Time to follow http://hf.co/tencent!

IndustryDGX agent

Time to follow http://hf.co/tencent! Genie3 generates videos. We generate 𝟯𝗗 𝘄𝗼𝗿𝗹𝗱𝘀 you can actually use. Launching tomorrow — Tencent #HYWorld 2.0, an engine-ready World Model🚀 This isn't a video. It

Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds

Model ReleasesDGX agent

arXiv:2604.09879v1 Announce Type: new Abstract: Deep neural networks for 3D point cloud understanding have achieved remarkable success in object classification and recognition, yet recent work shows t

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

Model ReleasesDGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning

Model ReleasesDGX agent

arXiv:2604.09850v1 Announce Type: new Abstract: Existing text-to-image diffusion models, while excelling at subject synthesis, exhibit a persistent foreground bias that treats the background as a pass

Trusted access for the next era of cyber defense

Model ReleasesDGX agent

OpenAI is expanding trusted access to its AI models for cybersecurity purposes, enabling vetted researchers, defenders, and organizations to leverage advanced AI capabilities for cyber defense applica

Weird Generalization is Weirdly Brittle

SafetyDGX agent

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

What and Where to Adapt: Structure-Semantics Co-Tuning for Machine Vision Compression via Synergistic Adapters

Model ReleasesDGX agent

arXiv:2604.10017v1 Announce Type: new Abstract: Parameter-efficient fine-tuning of pre-trained codecs is a promising direction in image compression for human and machine vision. While most existing wo

What Factors Affect LLMs and RLLMs in Financial Question Answering?

SafetyDGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

Model ReleasesDGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

XD-MAP: Cross-Modal Domain Adaptation via Semantic Parametric Maps for Scalable Training Data Generation

ResearchDGX agent

arXiv:2601.14477v2 Announce Type: replace-cross Abstract: Until open-world foundation models match the performance of specialized approaches, deep learning systems remain dependent on task- and sensor

13 Apr 2026

4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation

Model ReleasesDGX agent

arXiv:2512.17012v3 Announce Type: replace Abstract: Despite advances in Multimodal LLMs (MLLMs), their ability to reason over 3D structures and temporal dynamics remains limited, constrained by weak 4

Across the Levels of Analysis: Explaining Predictive Processing in Humans Requires More Than Machine-Estimated Probabilities

ResearchDGX agent

arXiv:2604.09466v1 Announce Type: new Abstract: Under the lens of Marr's levels of analysis, we critique and extend two claims about language models (LMs) and language processing: first, that predicti

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

Model ReleasesDGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

Automated Instruction Revision (AIR): A Structured Comparison of Task Adaptation Strategies for LLM

Model ReleasesDGX agent

arXiv:2604.09418v1 Announce Type: new Abstract: This paper studies Automated Instruction Revision (AIR), a rule-induction-based method for adapting large language models (LLMs) to downstream tasks usi

Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine

SafetyDGX agent

arXiv:2603.06665v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) often benefit from chain-of-thought (CoT) prompting in general domains, yet its efficacy in medical vision

Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding

Model ReleasesDGX agent

arXiv:2511.23071v2 Announce Type: replace-cross Abstract: Reading scene text, that is, text appearing in images, has numerous application areas, including assistive technology, search, and e-commerce.

Boosting Brain-inspired Path Integration Efficiency via Learning-based Replication of Continuous Attractor Neurodynamics

Model ReleasesDGX agent

arXiv:2511.17687v2 Announce Type: replace Abstract: The brain's Path Integration (PI) mechanism offers substantial guidance and inspiration for Brain-Inspired Navigation (BIN). However, the PI capabil

CAD 100K: A Comprehensive Multi-Task Dataset for Car Related Visual Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.09023v1 Announce Type: new Abstract: Multi-task visual anomaly detection is critical for car-related manufacturing quality assessment. However, existing methods remain task-specific, hinder

DeFakeQ: Enabling Real-Time Deepfake Detection on Edge Devices via Adaptive Bidirectional Quantization

Model ReleasesDGX agent

arXiv:2604.08847v1 Announce Type: new Abstract: Deepfake detection has become a fundamental component of modern media forensics. Despite significant progress in detection accuracy, most existing metho

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs

SafetyDGX agent

arXiv:2604.08846v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have been shown to be vulnerable to malicious queries that can elicit unsafe responses. Recent work uses prom

Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy

Model ReleasesDGX agent

arXiv:2604.09051v1 Announce Type: new Abstract: Fine-grained action segmentation during renorrhaphy in robot-assisted partial nephrectomy requires frame-level recognition of visually similar suturing

FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval

Model ReleasesDGX agent

arXiv:2604.09114v1 Announce Type: new Abstract: Composed image retrieval (CIR) aims to retrieve a target image that depicts a reference image modified by a textual description. While recent vision-lan

← Previous
1…432433434435436…1060
Next →