AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Applications

A multimodal and temporal foundation model for virtual patient representations at healthcare system scale

DGX agent

arXiv:2604.18570v1 Announce Type: cross Abstract: Modern medicine generates vast multimodal data across siloed systems, yet no existing model integrates the full breadth and temporal depth of the clin

applicationsarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

A Rapid Deployment Pipeline for Autonomous Humanoid Grasping Based on Foundation Models

DGX agent

arXiv:2604.17258v1 Announce Type: new Abstract: Deploying a humanoid robot to manipulate a new object has traditionally required one to two days of effort: data collection, manual annotation, 3D model

agentsarxiv-cs-ro
21 Apr 2026
Model Releases

Adversarial Humanities Benchmark: Results on Stylistic Robustness in Frontier Model Safety

DGX agent

arXiv:2604.18487v1 Announce Type: new Abstract: The Adversarial Humanities Benchmark (AHB) evaluates whether model safety refusals survive a shift away from familiar harmful prompt forms. Starting fro

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Aligning Language Models for Lyric-to-Melody Generation with Rule-Based Musical Constraints

DGX agent

arXiv:2604.18489v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in lyric-to-melody generation, but models trained with Supervised Fine-Tuning (SFT) often produce musically

safetyarxiv-cs-cl
21 Apr 2026
Applications

Appearance-free Action Recognition: Zero-shot Generalization in Humans and a Two-Pathway Model

DGX agent

arXiv:2604.16675v1 Announce Type: new Abstract: Action recognition is a fundamental ability for social species. Yet, its underlying computations are not well understood. Classical psychophysical studi

applicationsarxiv-cs-cv
21 Apr 2026
Research

Applications of deep generative models to DNA reaction kinetics and to cryogenic electron microscopy

DGX agent

arXiv:2604.16851v1 Announce Type: cross Abstract: This dissertation explores how deep generative models can advance the analysis of challenging biological problems by integrating domain knowledge with

researcharxiv-cs-cv
21 Apr 2026
Safety

Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models

DGX agent

arXiv:2604.18187v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have made significant progress in audio understanding, yet they primarily operate as perception-and-answer systems

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Auto-encoder model for faster generation of effective one-body gravitational waveform approximations

DGX agent

arXiv:2511.12642v2 Announce Type: replace-cross Abstract: Upgrades to current gravitational wave detectors for the next observation run and upcoming third-generation observatories, like the Einstein t

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Beyond the Failures: Rethinking Foundation Models in Pathology

DGX agent

arXiv:2510.23807v5 Announce Type: replace-cross Abstract: Despite their successes in vision and language, foundation models have stumbled in pathology, revealing low accuracy, instability, and heavy c

researcharxiv-cs-cv
21 Apr 2026
Model Releases

CaTS-Bench: Can Language Models Describe Time Series?

DGX agent

arXiv:2509.20823v5 Announce Type: replace-cross Abstract: Time series captioning, the task of describing time series in natural language, requires numeric and temporal reasoning, trend interpretation,

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

DifFoundMAD: Foundation Models meet Differential Morphing Attack Detection

DGX agent

arXiv:2604.17961v1 Announce Type: new Abstract: In this work, we introduce DifFoundMAD, a parameter-efficient D-MAD framework that exploits the generalisation capabilities of vision foundation models

model-releasesarxiv-cs-cv
21 Apr 2026
Tutorials

Dissipative Latent Residual Physics-Informed Neural Networks for Modeling and Identification of Electromechanical Systems

DGX agent

arXiv:2604.18277v1 Announce Type: new Abstract: Accurate dynamical modeling is essential for simulation and control of embodied systems, yet first-principles models of electromechanical systems often

tutorialsarxiv-cs-lg
21 Apr 2026
Research

Does AI See like Art Historians? Interpreting How Vision Language Models Recognize Artistic Style

DGX agent

arXiv:2603.11024v2 Announce Type: replace Abstract: VLMs have become increasingly proficient at a range of computer vision tasks, such as visual question answering and object detection. This includes

researcharxiv-cs-cv
21 Apr 2026
Research

Dual-End Consistency Model

DGX agent

arXiv:2602.10764v2 Announce Type: replace Abstract: The slow iterative sampling nature remains a major bottleneck for the practical deployment of diffusion and flow-based generative models. While cons

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization

DGX agent

arXiv:2604.17051v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated excellent performance in general language understanding, generation and other tasks. However, when fine-t

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Embedding Arithmetic: A Lightweight, Tuning-Free Framework for Post-hoc Bias Mitigation in Text-to-Image Models

DGX agent

arXiv:2604.18167v1 Announce Type: new Abstract: Modern text-to-image (T2I) models amplify harmful societal biases, challenging their ethical deployment. We introduce an inference-time method that reli

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning

DGX agent

arXiv:2412.02904v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized the field of natural language processing with their impressive reasoning and question-answering cap

researcharxiv-cs-cl
21 Apr 2026
Research

Geometry-Guided 3D Visual Token Pruning for Video-Language Models

DGX agent

arXiv:2604.18260v1 Announce Type: new Abstract: Multimodal large language models have demonstrated remarkable capabilities in 2D vision, motivating their extension to 3D scene understanding. Recent st

researcharxiv-cs-cv
21 Apr 2026
Applications

How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models

DGX agent

arXiv:2510.02370v3 Announce Type: replace Abstract: Large language models leverage both parametric knowledge acquired during pretraining and in-context knowledge provided at inference time. Crucially,

applicationsarxiv-cs-cl
21 Apr 2026
Agents

Human Cognition in Machines: A Unified Perspective of World Models

DGX agent

arXiv:2604.16592v1 Announce Type: cross Abstract: This comprehensive report distinguishes prior works by the cognitive functions they innovate. Many works claim an almost 'human-like' cognitive capabi

agentsarxiv-cs-cv
21 Apr 2026
Model Releases

ICAT: Incident-Case-Grounded Adaptive Testing for Physical-Risk Prediction in Embodied World Models

DGX agent

arXiv:2604.16405v1 Announce Type: cross Abstract: Video-generative world models are increasingly used as neural simulators for embodied planning and policy learning, yet their ability to predict physi

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

DGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

Lizard: An Efficient Linearization Framework for Large Language Models

DGX agent

arXiv:2507.09025v4 Announce Type: replace Abstract: We propose Lizard, a linearization framework that transforms pretrained Transformer-based Large Language Models (LLMs) into subquadratic architectur

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MetaLint: Easy-to-Hard Generalization for Code Linting

DGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Neural Network-Based Score Estimation in Diffusion Models: Optimization and Generalization

DGX agent

arXiv:2401.15604v4 Announce Type: replace Abstract: Diffusion models have become a leading paradigm in generative AI, with score estimation via denoising score matching as a central component. While r

researcharxiv-cs-lg
21 Apr 2026
Research

Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering

DGX agent

arXiv:2508.14461v3 Announce Type: replace Abstract: While multi-step diffusion models have advanced both forward and inverse rendering, existing approaches often treat these problems independently, le

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Please refuse to answer me! Mitigating Over-Refusal in Large Language Models via Adaptive Contrastive Decoding

DGX agent

arXiv:2604.17132v1 Announce Type: new Abstract: Safety-aligned large language models (LLMs) often generate refusal responses to harmless queries due to the over-refusal problem. However, existing meth

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

PoliLegalLM: A Technical Report on a Large Language Model for Political and Legal Affairs

DGX agent

arXiv:2604.17543v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable success in general-domain tasks, yet their direct application to the legal domain remains challeng

safetyarxiv-cs-cl
21 Apr 2026
Research

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning

DGX agent

arXiv:2502.02871v2 Announce Type: replace Abstract: Scientific reasoning, the process through which humans apply logic, evidence, and critical thinking to explore and interpret scientific phenomena, i

researcharxiv-cs-cl
21 Apr 2026
Research

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards

DGX agent

arXiv:2604.17957v1 Announce Type: new Abstract: Process Reward Models (PRMs) have emerged as a powerful tool for providing step-level feedback when evaluating the reasoning of Large Language Models (L

researcharxiv-cs-cl
21 Apr 2026
Model Releases

ProfVLM: A lightweight video-language model for multi-view proficiency estimation

DGX agent

arXiv:2509.26278v4 Announce Type: replace-cross Abstract: Most existing approaches formulate action quality assessment and skill proficiency estimation as discriminative prediction tasks, typically pr

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Real-Time Visual Attribution Streaming in Thinking Model

DGX agent

arXiv:2604.16587v1 Announce Type: new Abstract: We present an amortized framework for real-time visual attribution streaming in multimodal thinking models. When these models generate code from a scree

researcharxiv-cs-cv
21 Apr 2026
Research

REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations

DGX agent

arXiv:2604.17289v1 Announce Type: new Abstract: Supervised fine-tuning of large language models relies on human-annotated data, yet annotation pipelines routinely involve multiple crowdworkers of hete

researcharxiv-cs-lg
21 Apr 2026
Safety

RemoteShield: Enable Robust Multimodal Large Language Models for Earth Observation

DGX agent

arXiv:2604.17243v1 Announce Type: new Abstract: A robust Multimodal Large Language Model (MLLM) for Earth Observation should maintain consistent interpretation and reasoning under realistic input vari

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

ReTraceQA: Evaluating Reasoning Traces of Small Language Models in Commonsense Question Answering

DGX agent

arXiv:2510.09351v2 Announce Type: replace Abstract: While Small Language Models (SLMs) have demonstrated promising performance on an increasingly wide array of commonsense reasoning benchmarks, curren

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

S-GRPO: Unified Post-Training for Large Vision-Language Models

DGX agent

arXiv:2604.16557v1 Announce Type: cross Abstract: Current post-training methodologies for adapting Large Vision-Language Models (LVLMs) generally fall into two paradigms: Supervised Fine-Tuning (SFT)

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models

DGX agent

arXiv:2604.17873v1 Announce Type: new Abstract: Video Large Language Models (Vid-LLMs) have demonstrated remarkable performance in video understanding tasks, yet their robustness under conversational

model-releasesarxiv-cs-cv
21 Apr 2026
Research

TGLF-WINN: Data-Efficient Deep Learning Surrogate for Turbulent Transport Modeling in Fusion

DGX agent

arXiv:2509.07024v2 Announce Type: replace-cross Abstract: The Trapped Gyro-Landau Fluid (TGLF) model provides fast, accurate predictions of turbulent transport in tokamaks, but whole device simulation

researcharxiv-cs-lg
21 Apr 2026
Safety

Training Language Models to Use Prolog as a Tool

DGX agent

arXiv:2512.07407v2 Announce Type: replace Abstract: Language models frequently produce plausible yet incorrect reasoning traces that are difficult to verify. We investigate fine-tuning models to use P

safetyarxiv-cs-cl
21 Apr 2026
Safety

UniComp: A Unified Evaluation of Large Language Model Compression via Pruning, Quantization and Distillation

DGX agent

arXiv:2602.09130v3 Announce Type: replace Abstract: Model compression is increasingly essential for deploying large language models (LLMs), yet existing comparative studies largely focus on pruning an

safetyarxiv-cs-lg
21 Apr 2026
Research

Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models

DGX agent

arXiv:2506.00065v2 Announce Type: replace Abstract: Multimodal language models (MLMs) increasingly demonstrate human-like communication, yet their use of everyday perspectival words remains poorly und

researcharxiv-cs-cl
21 Apr 2026
Research

Vision Language Models are Biased

DGX agent

arXiv:2505.23941v4 Announce Type: replace-cross Abstract: Large language models (LLMs) memorize a vast amount of prior knowledge from the Internet that helps them on downstream tasks but also may noto

researcharxiv-cs-cv
21 Apr 2026
Safety

Where Do Self-Supervised Speech Models Become Unfair?

DGX agent

arXiv:2604.18249v1 Announce Type: new Abstract: Speech encoder models are known to model members of some speaker groups (SGs) better than others. However, there has been little work in establishing wh

safetyarxiv-cs-cl
21 Apr 2026
Research

LLMbench: A Comparative Close Reading Workbench for Large Language Models

DGX agent

arXiv:2604.15508v1 Announce Type: cross Abstract: LLMbench is a browser-based workbench for the comparative close reading of large language model (LLM) outputs. Where existing tools for LLM comparison

researcharxiv-cs-ai
20 Apr 2026
Research

Mechanisms of Prompt-Induced Hallucination in Vision-Language Models

DGX agent

arXiv:2601.05201v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) are highly capable, yet often hallucinate by favoring textual prompts over visual evidence. We study this

researcharxiv-cs-ai
20 Apr 2026
Model Releases

MM-Telco: Benchmarks and Multimodal Large Language Models for Telecom Applications

DGX agent

arXiv:2511.13131v2 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for automating complex reasoning and decision-making tasks. In telecommunications, they

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

MTR-DuplexBench: Towards a Comprehensive Evaluation of Multi-Round Conversations for Full-Duplex Speech Language Models

DGX agent

arXiv:2511.10262v3 Announce Type: replace-cross Abstract: Full-Duplex Speech Language Models (FD-SLMs) enable real-time, overlapping conversational interactions, offering a more dynamic user experienc

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

Security Threat Modeling for Emerging AI-Agent Protocols: A Comparative Analysis of MCP, A2A, Agora, and ANP

DGX agent

arXiv:2602.11327v2 Announce Type: replace-cross Abstract: The rapid development of the AI agent communication protocols, including the Model Context Protocol (MCP), Agent2Agent (A2A), Agora, and Agent

agentsarxiv-cs-ai
20 Apr 2026
← Previous
1…7273747576…1030
Next →