AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Research

LMMs Meet Object-Centric Vision: Understanding, Segmentation, Editing and Generation

DGX agent

arXiv:2604.11789v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have achieved remarkable progress in general-purpose vision--language understanding, yet they remain limited in tasks req

researcharxiv-cs-cv
14 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

LVSum: A Benchmark for Timestamp-Aware Long Video Summarization

DGX agent

arXiv:2604.10024v1 Announce Type: cross Abstract: Long video summarization presents significant challenges for current multimodal large language models (MLLMs), particularly in maintaining temporal fi

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

M2-Verify: A Large-Scale Multidomain Benchmark for Checking Multimodal Claim Consistency

DGX agent

arXiv:2604.01306v2 Announce Type: replace Abstract: Evaluating scientific arguments requires assessing the strict consistency between a claim and its underlying multimodal evidence. However, existing

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Mitigating Privacy Risk via Forget Set-Free Unlearning

DGX agent

arXiv:2604.10636v1 Announce Type: new Abstract: Training machine learning models requires the storage of large datasets, which often contain sensitive or private data. Storing data is associated with

researcharxiv-cs-lg
14 Apr 2026
Safety

MM-LIMA: Less Is More for Alignment in Multi-Modal Datasets

DGX agent

arXiv:2308.12067v3 Announce Type: replace-cross Abstract: Multimodal large language models are typically trained in two stages: first pre-training on image-text pairs, and then fine-tuning using super

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization

DGX agent

arXiv:2604.11259v1 Announce Type: new Abstract: Mobile GUI agents powered by Multimodal Large Language Models (MLLMs) can execute complex tasks on mobile devices. Despite this progress, most existing

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Muon^2: Boosting Muon via Adaptive Second-Moment Preconditioning

DGX agent

arXiv:2604.09967v1 Announce Type: cross Abstract: Muon has emerged as a promising optimizer for large-scale foundation model pre-training by exploiting the matrix structure of neural network updates t

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Nano-EmoX: Unifying Multimodal Emotional Intelligence from Perception to Empathy

DGX agent

arXiv:2603.02123v3 Announce Type: replace Abstract: The development of affective multimodal language models (MLMs) has long been constrained by a gap between low-level perception and high-level intera

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Natural Gradient Gaussian Approximation Filter on Lie Groups for Robot State Estimation

DGX agent

arXiv:2604.10057v1 Announce Type: new Abstract: Accurate state estimation for robotic systems evolving on Lie group manifolds, such as legged robots, is a prerequisite for achieving agile control. How

model-releasesarxiv-cs-ro
14 Apr 2026
Research

NetworkNet: A Deep Neural Network Approach for Random Networks with Sparse Nodal Attributes and Complex Nodal Heterogeneity

DGX agent

arXiv:2604.11673v1 Announce Type: cross Abstract: Heterogeneous network data with rich nodal information become increasingly prevalent across multidisciplinary research, yet accurately modeling comple

researcharxiv-cs-ai
14 Apr 2026
Hardware

NVIDIA Ising Introduces AI-Powered Workflows to Build Fault-Tolerant Quantum Systems

DGX agent

NVIDIA Ising is the world's first family of open-source quantum AI models, designed to help researchers and enterprises build quantum processors capable of running useful applications. The family span

hardwarenvidia-developer
14 Apr 2026
Model Releases

ODUTQA-MDC: A Task for Open-Domain Underspecified Tabular QA with Multi-turn Dialogue-based Clarification

DGX agent

arXiv:2604.10159v1 Announce Type: new Abstract: The advancement of large language models (LLMs) has enhanced tabular question answering (Tabular QA), yet they struggle with open-domain queries exhibit

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Online Learning-Enhanced High Order Adaptive Safety Control

DGX agent

arXiv:2511.19651v2 Announce Type: replace Abstract: Control barrier functions (CBFs) are an effective model-based tool to formally certify the safety of a system. With the growing complexity of modern

safetyarxiv-cs-ro
14 Apr 2026
Model Releases

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents dep…

DGX agent

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents deploy has easy configs to let users customize their harness and

model-releasesharrison-chase--x
14 Apr 2026
Applications

Pacing Opinion Polarization via Graph Reinforcement Learning

DGX agent

arXiv:2602.23390v2 Announce Type: replace-cross Abstract: Opinion polarization moderation has been studied mainly as an analytical optimization problem under the Friedkin Johnson FJ model, where inter

applicationsarxiv-cs-lg
14 Apr 2026
Research

PanoSAMic: Panoramic Image Segmentation from SAM Feature Encoding and Dual View Fusion

DGX agent

arXiv:2601.07447v2 Announce Type: replace Abstract: Existing image foundation models are not optimized for spherical images having been trained primarily on perspective images. PanoSAMic integrates th

researcharxiv-cs-cv
14 Apr 2026
Model Releases

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

DGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

model-releasesarxiv-cs-ai
14 Apr 2026
Research

PAS: Estimating the target accuracy before domain adaptation

DGX agent

arXiv:2604.09863v1 Announce Type: cross Abstract: The goal of domain adaptation is to make predictions for unlabeled samples from a target domain with the help of labeled samples from a different but

researcharxiv-cs-ai
14 Apr 2026
Research

Poisoning with A Pill: Circumventing Detection in Federated Learning

DGX agent

arXiv:2407.15389v2 Announce Type: replace Abstract: Without direct access to the client's data, federated learning (FL) is well-known for its unique strength in data privacy protection among existing

researcharxiv-cs-lg
14 Apr 2026
Local Ai

Post-Processing Methods for Improving Accuracy in MRI Inpainting

DGX agent

arXiv:2510.15282v2 Announce Type: replace-cross Abstract: Magnetic Resonance Imaging (MRI) is the primary imaging modality used in the diagnosis, assessment, and treatment planning for brain pathologi

local-aiarxiv-cs-ai
14 Apr 2026
Model Releases

Precision Synthesis of Multi-Tracer PET via VLM-Modulated Rectified Flow for Stratifying Mild Cognitive Impairment

DGX agent

arXiv:2604.11176v1 Announce Type: new Abstract: The biological definition of Alzheimer's disease (AD) relies on multi-modal neuroimaging, yet the clinical utility of positron emission tomography (PET)

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

PRIX: Learning to Plan from Raw Pixels for End-to-End Autonomous Driving

DGX agent

arXiv:2507.17596v3 Announce Type: replace-cross Abstract: While end-to-end autonomous driving models show promising results, their practical deployment is often hindered by large model sizes, a relian

agentsarxiv-cs-ai
14 Apr 2026
Tutorials

Probabilistic Prediction of Neural Dynamics via Autoregressive Flow Matching

DGX agent

arXiv:2604.11178v1 Announce Type: cross Abstract: Forecasting neural activity in response to naturalistic stimuli remains a key challenge for understanding brain dynamics and enabling downstream neuro

tutorialsarxiv-cs-lg
14 Apr 2026
Local Ai

Psychological Concept Neurons: Can Neural Control Bias Probing and Shift Generation in LLMs?

DGX agent

arXiv:2604.11802v1 Announce Type: new Abstract: Using psychological constructs such as the Big Five, large language models (LLMs) can imitate specific personality profiles and predict a user's persona

local-aiarxiv-cs-cl
14 Apr 2026
Model Releases

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

DGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought

DGX agent

arXiv:2506.16796v4 Announce Type: replace Abstract: Real-World Image Super-Resolution is one of the most challenging task in image restoration. However, existing methods struggle with an accurate unde

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

Reasoning as Gradient: Scaling MLE Agents Beyond Tree Search

DGX agent

arXiv:2603.01692v3 Announce Type: replace-cross Abstract: LLM-based agents for machine learning engineering (MLE) predominantly rely on tree search, a form of gradient-free optimization that uses scal

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization

DGX agent

arXiv:2604.10520v1 Announce Type: cross Abstract: As Large Language Models (LLMs) have become capable of generating long and descriptive code summaries, accurate and reliable evaluation of factual con

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

DGX agent

arXiv:2604.11554v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language mode

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

DGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RL makes MLLMs see better than SFT

DGX agent

arXiv:2510.16333v2 Announce Type: replace Abstract: A dominant assumption in Multimodal Language Model (MLLM) research is that its performance is largely inherited from the LLM backbone, given its imm

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Robust Fair Disease Diagnosis in CT Images

DGX agent

arXiv:2604.09710v1 Announce Type: new Abstract: Automated diagnosis from chest CT has improved considerably with deep learning, but models trained on skewed datasets tend to perform unevenly across pa

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

DGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

local-aiarxiv-cs-ai
14 Apr 2026
Model Releases

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

DGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

DGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding

DGX agent

arXiv:2604.11122v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated immense potential in Earth observation. However, the massive visual tokens generated when p

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence

DGX agent

arXiv:2604.10404v1 Announce Type: cross Abstract: Edge-based multimodal medical monitoring requires models that balance diagnostic accuracy with severe energy constraints. Continuous acquisition of EC

safetyarxiv-cs-lg
14 Apr 2026
Model Releases

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

DGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Steered LLM Activations are Non-Surjective

DGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

DGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

model-releasesarxiv-cs-cv
14 Apr 2026
Research

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

DGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

researcharxiv-cs-ai
14 Apr 2026
Research

Teaching Robots to Interpret Social Interactions through Lexically-guided Dynamic Graph Learning

DGX agent

arXiv:2604.10895v1 Announce Type: cross Abstract: For a robot to be called socially intelligent, it must be able to infer users internal states from their current behaviour, predict the users future b

researcharxiv-cs-ro
14 Apr 2026
Model Releases

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

DGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation

DGX agent

arXiv:2510.15552v3 Announce Type: replace-cross Abstract: Large language models (LLMs) still struggle with multi-hop reasoning over knowledge-graphs (KGs), and we identify a previously overlooked stru

model-releasesarxiv-cs-ai
14 Apr 2026
Industry

Time to follow http://hf.co/tencent!

DGX agent

Time to follow http://hf.co/tencent! Genie3 generates videos. We generate 𝟯𝗗 𝘄𝗼𝗿𝗹𝗱𝘀 you can actually use. Launching tomorrow — Tencent #HYWorld 2.0, an engine-ready World Model🚀 This isn't a video. It

industryclem-delangue--x
14 Apr 2026
Model Releases

Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds

DGX agent

arXiv:2604.09879v1 Announce Type: new Abstract: Deep neural networks for 3D point cloud understanding have achieved remarkable success in object classification and recognition, yet recent work shows t

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

DGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning

DGX agent

arXiv:2604.09850v1 Announce Type: new Abstract: Existing text-to-image diffusion models, while excelling at subject synthesis, exhibit a persistent foreground bias that treats the background as a pass

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…560561562563564…1371
Next →