AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models

DGX agent

arXiv:2604.12995v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into real-world decision-making, including in the domain of public policy. Yet, their ability t

model-releasesarxiv-cs-cl
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

A Tale of Two Temperatures: Simple, Efficient, and Diverse Sampling from Diffusion Language Models

DGX agent

arXiv:2604.09921v1 Announce Type: new Abstract: Much work has been done on designing fast and accurate sampling for diffusion language models (dLLMs). However, these efforts have largely focused on th

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

Assessing the Pedagogical Readiness of Large Language Models as AI Tutors in Low-Resource Contexts: A Case Study of Nepal's K-10 Curriculum

DGX agent

arXiv:2604.09619v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into educational ecosystems promises to democratize access to personalized tutoring, yet the readiness

model-releasesarxiv-cs-ai
14 Apr 2026
Research

bacpipe: a Python package to make bioacoustic deep learning models accessible

DGX agent

arXiv:2604.11560v1 Announce Type: cross Abstract: 1. Natural sounds have been recorded for millions of hours over the previous decades using passive acoustic monitoring. Improvements in deep learning

researcharxiv-cs-ai
14 Apr 2026
Model Releases

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

DGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Comparative Analysis of Large Language Models in Healthcare

DGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

model-releasesarxiv-cs-cl
14 Apr 2026
Research

DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models

DGX agent

arXiv:2604.11572v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) have demonstrated strong potential for embodied AI, yet their deployment on resource-limited robots remains challen

researcharxiv-cs-ro
14 Apr 2026
Model Releases

Enhancing Multimodal Large Language Models for Ancient Chinese Character Evolution Analysis via Glyph-Driven Fine-Tuning

DGX agent

arXiv:2604.11299v1 Announce Type: cross Abstract: In recent years, rapid advances in Multimodal Large Language Models (MLLMs) have increasingly stimulated research on ancient Chinese scripts. As the e

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Environmental Footprint of GenAI Research: Insights from the Moshi Foundation Model

DGX agent

arXiv:2604.11154v1 Announce Type: new Abstract: New multi-modal large language models (MLLMs) are continuously being trained and deployed, following rapid development cycles. This generative AI frenzy

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Evaluating Reliability Gaps in Large Language Model Safety via Repeated Prompt Sampling

DGX agent

arXiv:2604.09606v1 Announce Type: new Abstract: Traditional benchmarks for large language models (LLMs), such as HELM and AIR-BENCH, primarily assess safety risk through breadth-oriented evaluation ac

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

GeoArena: Evaluating Open-World Geographic Reasoning in Large Vision-Language Models

DGX agent

arXiv:2509.04334v4 Announce Type: replace Abstract: Geographic reasoning is a fundamental cognitive capability that requires models to infer plausible locations by synthesizing visual evidence with sp

model-releasesarxiv-cs-cv
14 Apr 2026
Research

GS4City: Hierarchical Semantic Gaussian Splatting via City-Model Priors

DGX agent

arXiv:2604.11401v1 Announce Type: new Abstract: Recent semantic 3D Gaussian Splatting (3DGS) methods primarily rely on 2D foundation models, often yielding ambiguous boundaries and limited support for

researcharxiv-cs-cv
14 Apr 2026
Model Releases

How Robust Are Large Language Models for Clinical Numeracy? An Empirical Study on Numerical Reasoning Abilities in Clinical Contexts

DGX agent

arXiv:2604.11133v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being explored for clinical question answering and decision support, yet safe deployment critically requir

model-releasesarxiv-cs-cl
14 Apr 2026
Applications

Inferring Dynamic Physical Properties from Video Foundation Models

DGX agent

arXiv:2510.02311v2 Announce Type: replace Abstract: We study the task of predicting dynamic physical properties from videos. More specifically, we consider physical properties that require temporal in

applicationsarxiv-cs-cv
14 Apr 2026
Safety

Influencing Humans to Conform to Preference Models for RLHF

DGX agent

arXiv:2501.06416v3 Announce Type: replace-cross Abstract: Designing a reinforcement learning from human feedback (RLHF) algorithm to approximate a human's unobservable reward function requires assumin

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Is There Knowledge Left to Extract? Evidence of Fragility in Medically Fine-Tuned Vision-Language Models

DGX agent

arXiv:2604.09841v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly adapted through domain-specific fine-tuning, yet it remains unclear whether this improves reasoning bey

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion

DGX agent

arXiv:2604.10326v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak attacks -- inputs designed to bypass safety mechanisms and elicit harmful responses -- despite ad

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

LaMI: Augmenting Large Language Models via Late Multi-Image Fusion

DGX agent

arXiv:2406.13621v2 Announce Type: replace Abstract: Commonsense reasoning often requires both textual and visual knowledge, yet Large Language Models (LLMs) trained solely on text lack visual groundin

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Latent Structure of Affective Representations in Large Language Models

DGX agent

arXiv:2604.07382v2 Announce Type: replace-cross Abstract: The geometric structure of latent representations in large language models (LLMs) is an active area of research, driven in part by its implica

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models

DGX agent

arXiv:2511.18373v2 Announce Type: replace Abstract: Vision Language Models (VLMs) perform well on standard video tasks but struggle with physics-related reasoning involving motion dynamics and spatial

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model

DGX agent

arXiv:2505.23606v4 Announce Type: replace-cross Abstract: Unified generation models aim to handle diverse tasks across modalities -- such as text generation, image generation, and vision-language reas

researcharxiv-cs-cv
14 Apr 2026
Model Releases

NovBench: Evaluating Large Language Models on Academic Paper Novelty Assessment

DGX agent

arXiv:2604.11543v1 Announce Type: cross Abstract: Novelty is a core requirement in academic publishing and a central focus of peer review, yet the growing volume of submissions has placed increasing p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

On Harnessing Idle Compute at the Edge for Foundation Model Training

DGX agent

arXiv:2512.22142v2 Announce Type: replace-cross Abstract: The foundation-model ecosystem remains highly centralized because training requires immense compute resources and is therefore largely limited

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

PSF-Med: Measuring and Explaining Paraphrase Sensitivity in Medical Vision Language Models

DGX agent

arXiv:2602.21428v2 Announce Type: replace Abstract: Medical Vision Language Models (VLMs) can change their answers when clinicians rephrase the same question, a failure mode that threatens deployment

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging

DGX agent

arXiv:2604.11399v1 Announce Type: cross Abstract: Multimodal adaptation equips large language models (LLMs) with perceptual capabilities, but often weakens the reasoning ability inherited from languag

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Retrieval-Augmented Large Language Models for Evidence-Informed Guidance on Cannabidiol Use in Older Adults

DGX agent

arXiv:2604.09548v1 Announce Type: cross Abstract: Older adults commonly experience chronic conditions such as pain and sleep disturbances and may consider cannabidiol for symptom management. Safe use

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Seeing Through the Tool: A Controlled Benchmark for Occlusion Robustness in Foundation Segmentation Models

DGX agent

arXiv:2604.11711v1 Announce Type: new Abstract: Occlusion, where target structures are partially hidden by surgical instruments or overlapping tissues, remains a critical yet underexplored challenge f

model-releasesarxiv-cs-cv
14 Apr 2026
Research

ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values

DGX agent

arXiv:2604.11200v1 Announce Type: cross Abstract: Changes in input distribution can induce shifts in the average predictions of machine learning models. Such prediction shifts may impact downstream bu

researcharxiv-cs-ai
14 Apr 2026
Model Releases

SODA: Semi On-Policy Black-Box Distillation for Large Language Models

DGX agent

arXiv:2604.03873v2 Announce Type: replace-cross Abstract: Black-box knowledge distillation for large language models presents a strict trade-off. Simple off-policy methods (e.g., sequence-level knowle

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Text-to-Image Models and Their Representation of People from Different Nationalities Engaging in Activities

DGX agent

arXiv:2504.06313v5 Announce Type: replace Abstract: This paper investigates how popular text-to-image (T2I) models, DALL-E 3 and Gemini 3 Pro Preview, depict people from 206 nationalities when prompte

model-releasesarxiv-cs-cv
14 Apr 2026
Applications

Tuning Language Models for Robust Prediction of Diverse User Behaviors

DGX agent

arXiv:2505.17682v2 Announce Type: replace-cross Abstract: Predicting user behavior is essential for intelligent assistant services, yet deep learning models often struggle to capture long-tailed behav

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models

DGX agent

arXiv:2604.10079v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the standard approach for adapting large language models (LLMs) to downstream tasks. However, we observe a persistent fa

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Breaking Block Boundaries: Anchor-based History-stable Decoding for Diffusion Large Language Models

DGX agent

arXiv:2604.08964v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have recently become a promising alternative to autoregressive large language models (ARMs). Semi-autoregressive

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Where Vision Becomes Text: Locating the OCR Routing Bottleneck in Vision-Language Models

DGX agent

arXiv:2602.22918v2 Announce Type: replace Abstract: Vision-language models (VLMs) can read text from images, but where does this optical character recognition (OCR) information enter the language proc

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

CAMO: A Class-Aware Minority-Optimized Ensemble for Robust Language Model Evaluation on Imbalanced Data

DGX agent

arXiv:2604.07583v1 Announce Type: new Abstract: Real-world categorization is severely hampered by class imbalance because traditional ensembles favor majority classes, which lowers minority performanc

model-releasesarxiv-cs-cl
10 Apr 2026
Research

DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification

DGX agent

arXiv:2604.07166v1 Announce Type: cross Abstract: Although visual foundation models like DINOv2 provide state-of-the-art performance as feature extractors, their complex, high-dimensional representati

researcharxiv-cs-lg
10 Apr 2026
Research

DMin: Scalable Training Data Influence Estimation for Diffusion Models

DGX agent

arXiv:2412.08637v4 Announce Type: replace Abstract: Identifying the training data samples that most influence a generated image is a critical task in understanding diffusion models (DMs), yet existing

researcharxiv-cs-cv
10 Apr 2026
Model Releases

MF-GLaM: A multifidelity stochastic emulator using generalized lambda models

DGX agent

arXiv:2507.10303v2 Announce Type: replace-cross Abstract: Stochastic simulators exhibit intrinsic stochasticity due to unobservable, uncontrollable, or unmodeled input variables, resulting in random o

model-releasesarxiv-cs-lg
10 Apr 2026
Safety

MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models

DGX agent

arXiv:2604.07991v1 Announce Type: new Abstract: Recent advances in world models have demonstrated strong capabilities in simulating physical reality, making them an increasingly important foundation f

safetyarxiv-cs-cv
10 Apr 2026
Applications

Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism

DGX agent

arXiv:2510.26083v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at general language tasks but struggle in specialized domains. Specialized Generalist Models (SGMs) address

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

On Emotion-Sensitive Decision Making of Small Language Model Agents

DGX agent

arXiv:2604.06562v1 Announce Type: new Abstract: Small language models (SLM) are increasingly used as interactive decision-making agents, yet most decision-oriented evaluations ignore emotion as a caus

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Self-Preference Bias in Rubric-Based Evaluation of Large Language Models

DGX agent

arXiv:2604.06996v1 Announce Type: cross Abstract: LLM-as-a-judge has become the de facto approach for evaluating LLM outputs. However, judges are known to exhibit self-preference bias (SPB): they tend

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Spatio-Temporal Grounding of Large Language Models from Perception Streams

DGX agent

arXiv:2604.07592v1 Announce Type: new Abstract: Embodied-AI agents must reason about how objects move and interact in 3-D space over time, yet existing smaller frontier Large Language Models (LLMs) st

model-releasesarxiv-cs-ro
10 Apr 2026
Model Releases

SUPERGLASSES: Benchmarking Vision Language Models as Intelligent Agents for AI Smart Glasses

DGX agent

arXiv:2602.22683v2 Announce Type: replace Abstract: The rapid advancement of AI-powered smart glasses-one of the hottest wearable devices-has unlocked new frontiers for multimodal interaction, with Vi

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning

DGX agent

arXiv:2604.06427v1 Announce Type: cross Abstract: The viability of chain-of-thought (CoT) monitoring hinges on models being unable to reason effectively in their latent representations. Yet little is

model-releasesarxiv-cs-ai
10 Apr 2026
Research

The Detection-Extraction Gap: Models Know the Answer Before They Can Say It

DGX agent

arXiv:2604.06613v2 Announce Type: cross Abstract: Modern reasoning models continue generating long after the answer is already determined. Across five model configurations, two families, and three ben

researcharxiv-cs-ai
10 Apr 2026
Model Releases

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval

DGX agent

arXiv:2512.08410v2 Announce Type: replace Abstract: Due to excessive memory overhead, most Multimodal Large Language Models (MLLMs) can only process videos of limited frames. In this paper, we propose

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

DGX agent

arXiv:2604.08362v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) has illuminated the potential for a general-purpose user simulator. However, existing benchmarks remain co

model-releasesarxiv-cs-cl
10 Apr 2026
← Previous
1…5556575859…1021
Next →