AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

TCM-Serve: Modality-aware Scheduling for Multimodal Large Language Model Inference

DGX agent

arXiv:2603.26498v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) power platforms like ChatGPT, Gemini, and Copilot, enabling richer interactions with text, images, an

model-releasesarxiv-cs-ai
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

The Shape of Beliefs: Geometry, Dynamics, and Interventions along Representation Manifolds of Language Models' Posteriors

DGX agent

arXiv:2602.02315v2 Announce Type: replace Abstract: Large language models (LLMs) form implicit beliefs (posteriors over latent variables) from prompts, but we lack a mechanistic account of how these b

model-releasesarxiv-cs-cl
7 May 2026
Safety

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models

DGX agent

arXiv:2605.04874v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has proven to be an effective solution for mitigating hallucination in Multimodal Large Language Models (MLLMs) b

safetyarxiv-cs-cl
7 May 2026
Local Ai

A Few-Step Generative Model on Cumulative Flow Maps

DGX agent

arXiv:2605.03623v1 Announce Type: new Abstract: We propose a unified, few-step generative modeling framework based on cumulative flow maps for long-range transport in probability space, inspired by fl

local-aiarxiv-cs-lg
6 May 2026
Applications

A Unified Framework for Tabular Generative Modeling: Loss Functions, Benchmarks, and Improved Multi-objective Bayesian Optimization Approaches

DGX agent

arXiv:2405.16971v2 Announce Type: replace Abstract: Deep learning (DL) models require extensive data to achieve strong performance and generalization. Deep generative models (DGMs) offer a solution by

applicationsarxiv-cs-lg
6 May 2026
Safety

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling

DGX agent

arXiv:2507.07982v2 Announce Type: replace Abstract: Videos inherently represent 2D projections of a dynamic 3D world. However, our analysis suggests that video diffusion models trained solely on raw v

safetyarxiv-cs-cv
6 May 2026
Safety

GRPO-TTA: Test-Time Visual Tuning for Vision-Language Models via GRPO-Driven Reinforcement Learning

DGX agent

arXiv:2605.03403v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has recently shown strong performance in post-training large language models and vision-language models. It ra

safetyarxiv-cs-cv
6 May 2026
Model Releases

Hybrid Models for Natural Language Reasoning: The Case of Syllogistic Logic

DGX agent

arXiv:2510.09472v2 Announce Type: replace Abstract: Despite the remarkable progress in neural models, their ability to generalize, a cornerstone for applications such as logical reasoning, remains a c

model-releasesarxiv-cs-cl
6 May 2026
Applications

Joint Relational Database Generation via Graph-Conditional Diffusion Models

DGX agent

arXiv:2505.16527v2 Announce Type: replace Abstract: Building generative models for relational databases (RDBs) is important for many applications, such as privacy-preserving data release and augmentin

applicationsarxiv-cs-lg
6 May 2026
Safety

Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation

DGX agent

arXiv:2605.01790v1 Announce Type: cross Abstract: A common design pattern in high-quality music generation is to handle structure and fidelity in different representation spaces: a generator first mod

safetyarxiv-cs-ai
6 May 2026
Agents

Latent State Design for World Models under Sufficiency Constraints

DGX agent

arXiv:2605.01694v1 Announce Type: new Abstract: A world model matters to an agent only through the state it constructs. That state must preserve some information, discard other information, and suppor

agentsarxiv-cs-ai
6 May 2026
Model Releases

MHPR: Multidimensional Human Perception and Reasoning Benchmark for Large Vision-Languate Models

DGX agent

arXiv:2605.03485v1 Announce Type: new Abstract: Multidimensional human understanding is essential for real-world applications such as film analysis and virtual digital humans, yet current LVLM benchma

model-releasesarxiv-cs-cv
6 May 2026
Tutorials

SCPRM: A Schema-aware Cumulative Process Reward Model for Knowledge Graph Question Answering

DGX agent

arXiv:2605.02819v1 Announce Type: new Abstract: Large language models excel at complex reasoning, yet evaluating their intermediate steps remains challenging. Although process reward models provide st

tutorialsarxiv-cs-ai
6 May 2026
Local Ai

SHIELD: A Diverse Clinical Note Dataset and Distilled Small Language Models for Enterprise-Scale De-identification

DGX agent

arXiv:2605.03301v1 Announce Type: new Abstract: De-identification of clinical text remains essential for secondary use of electronic health records (EHRs), yet public benchmarks such as i2b2 2006/2014

local-aiarxiv-cs-cl
6 May 2026
Model Releases

Vibe Code Bench: Evaluating AI Models on End-to-End Web Application Development

DGX agent

arXiv:2603.04601v2 Announce Type: replace-cross Abstract: Code generation has emerged as one of AI's highest-impact use cases, yet existing benchmarks measure isolated tasks rather than the complete '

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning

DGX agent

arXiv:2605.02263v1 Announce Type: new Abstract: Recent diffusion large language models (dLLMs) have demonstrated both effectiveness and efficiency in reasoning via a block-based semi-autoregressive ge

model-releasesarxiv-cs-lg
5 May 2026
Research

Counting as a minimal probe of language model reliability

DGX agent

arXiv:2605.02028v1 Announce Type: new Abstract: Large language models perform strongly on benchmarks in mathematical reasoning, coding and document analysis, suggesting a broad ability to follow instr

researcharxiv-cs-cl
5 May 2026
Model Releases

Empowering Heterogeneous Graph Foundation Models via Decoupled Relation Alignment

DGX agent

arXiv:2605.00731v1 Announce Type: cross Abstract: While Graph Foundation Models (GFMs) have achieved remarkable success in homogeneous graphs, extending them to multi-domain heterogeneous graphs (MDHG

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

GaMMA: Towards Joint Global-Temporal Music Understanding in Large Multimodal Models

DGX agent

arXiv:2605.00371v1 Announce Type: cross Abstract: In this paper, we propose GaMMA, a state-of-the-art (SoTA) large multimodal model (LMM) designed to achieve comprehensive musical content understandin

model-releasesarxiv-cs-ai
5 May 2026
Tutorials

GenRecEdit: Adapting Model Editing for Generative Recommendation with Cold-Start Items

DGX agent

arXiv:2603.14259v2 Announce Type: replace-cross Abstract: Generative recommendation (GR) has shown strong potential for sequential recommendation in an end-to-end generation paradigm. However, existin

tutorialsarxiv-cs-ai
5 May 2026
Model Releases

Importance-Guided Basis Selection for Low-Rank Decomposition of Large Language Models

DGX agent

arXiv:2605.01627v1 Announce Type: new Abstract: Low-rank decomposition is a compelling approach for compressing large language models, but its effectiveness hinges on selecting which singular-vector b

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Molecular Representations for Large Language Models

DGX agent

arXiv:2605.01822v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being used to support scientific discovery. In chemistry, tasks such as reaction prediction and structure

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

OceanPile: A Large-Scale Multimodal Ocean Corpus for Foundation Models

DGX agent

arXiv:2605.00877v1 Announce Type: cross Abstract: The vast and underexplored ocean plays a critical role in regulating global climate and supporting marine biodiversity, yet artificial intelligence ha

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval

DGX agent

arXiv:2605.02623v1 Announce Type: new Abstract: Video Moment Retrieval (VMR) aims to localize temporal segments in videos that correspond to a natural language query, but typically assumes only a sing

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Shadow-Loom: Causal Reasoning over Graphical World Model of Narratives

DGX agent

arXiv:2605.02475v1 Announce Type: cross Abstract: Stories hold a reader's attention because they have causes, secrets, and consequences. Shadow-Loom is an experimental open-source framework that turns

model-releasesarxiv-cs-cl
5 May 2026
Research

Skipping the Zeros in Diffusion Models for Sparse Data Generation

DGX agent

arXiv:2605.01817v1 Announce Type: new Abstract: Diffusion models (DMs) excel on dense continuous data, but are not designed for sparse continuous data. They do not model exact zeros that represent the

researcharxiv-cs-lg
5 May 2026
Applications

SlimDiffSR: Toward Lightweight and Efficient Remote Sensing Image Super-Resolution via Diffusion Model Distillation

DGX agent

arXiv:2605.02198v1 Announce Type: new Abstract: Diffusion models have recently achieved remarkable performance in image super-resolution (SR), but their high computational cost limits practical deploy

applicationsarxiv-cs-cv
5 May 2026
Tutorials

SpecTM: Spectral Targeted Masking for Trustworthy Foundation Models

DGX agent

arXiv:2603.22097v2 Announce Type: replace-cross Abstract: Foundation models are now increasingly being developed for Earth observation (EO), yet they often rely on stochastic masking that do not expli

tutorialsarxiv-cs-lg
5 May 2026
Model Releases

SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?

DGX agent

arXiv:2605.01911v1 Announce Type: new Abstract: Purpose: Vision-language models (VLMs) have shown promising performance in surgical visual question answering (VQA). However, existing surgical VQA data

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

TF1-EN-3M: Three Million Synthetic Moral Fables for Training Small, Open Language Models

DGX agent

arXiv:2504.20605v2 Announce Type: replace Abstract: Moral stories are a time-tested vehicle for transmitting values, yet modern NLP lacks a large, structured corpus that couples coherent narratives wi

model-releasesarxiv-cs-cl
5 May 2026
Research

Validation of Whole-Slide Foundation Models for Image Retrieval in TCGA Data

DGX agent

arXiv:2605.00902v1 Announce Type: new Abstract: Foundation models are reshaping computational histopathology, yet their value for whole-slide image retrieval relative to strong patch-based and supervi

researcharxiv-cs-cv
5 May 2026
Applications

Being-H0.7: A Latent World-Action Model from Egocentric Videos

DGX agent

arXiv:2605.00078v1 Announce Type: cross Abstract: Visual-Language-Action models (VLAs) have advanced generalist robot control by mapping multimodal observations and language instructions directly to a

applicationsarxiv-cs-cv
4 May 2026
Research

Beyond Decodability: Reconstructing Language Model Representations with an Encoding Probe

DGX agent

arXiv:2605.00607v1 Announce Type: new Abstract: Probing is widely used to study which features can be decoded from language model representations. However, the common decoding probe approach has two l

researcharxiv-cs-cl
4 May 2026
Safety

Bias in Large Language Models: Origin, Evaluation, and Mitigation

DGX agent

arXiv:2411.10915v2 Announce Type: replace Abstract: Large Language Models (LLMs) have revolutionized natural language processing, but their susceptibility to biases poses significant challenges. This

safetyarxiv-cs-cl
4 May 2026
Safety

Can Small Language Models Handle Context-Summarized Multi-Turn Customer-Service QA? A Synthetic Data-Driven Comparative Evaluation

DGX agent

arXiv:2602.00665v3 Announce Type: replace Abstract: Customer-service question answering (QA) systems increasingly rely on conversational language understanding. While Large Language Models (LLMs) achi

safetyarxiv-cs-cl
4 May 2026
Research

CollaFuse: Collaborative Diffusion Models

DGX agent

arXiv:2406.14429v3 Announce Type: replace-cross Abstract: In the landscape of generative artificial intelligence, diffusion-based models have emerged as a promising method for generating synthetic ima

researcharxiv-cs-cv
4 May 2026
Model Releases

Differentiable Autoencoding Neural Operator for Interpretable and Integrable Latent Space Modeling

DGX agent

arXiv:2510.00233v2 Announce Type: replace Abstract: Scientific machine learning has enabled the extraction of physical insights and data-driven modeling of high-dimensional spatiotemporal data, yet ac

model-releasesarxiv-cs-lg
4 May 2026
Applications

Foundation Models for Discovery and Exploration in Chemical Space

DGX agent

arXiv:2510.18900v2 Announce Type: replace-cross Abstract: Accurate prediction of atomistic, thermodynamic, and kinetic properties from molecular structures underpins materials innovation. Existing com

applicationsarxiv-cs-lg
4 May 2026
Model Releases

Jailbreaking Vision-Language Models Through the Visual Modality

DGX agent

arXiv:2605.00583v1 Announce Type: new Abstract: The visual modality of vision-language models (VLMs) is an underexplored attack surface for bypassing safety alignment. We introduce four jailbreak atta

model-releasesarxiv-cs-cv
4 May 2026
Research

LandSegmenter: Towards a Flexible Foundation Model for Land Use and Land Cover Mapping

DGX agent

arXiv:2511.08156v2 Announce Type: replace Abstract: Land Use and Land Cover (LULC) mapping is a fundamental task in Earth Observation (EO). However, current LULC models are typically developed for a s

researcharxiv-cs-cv
4 May 2026
Model Releases

Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context

DGX agent

arXiv:2505.22003v2 Announce Type: replace Abstract: In India, access to legal assistance for the general public has been observed to have a critical gap, as many citizens are not able to take full adv

model-releasesarxiv-cs-cl
4 May 2026
Research

LLM DNA: Tracing Model Evolution via Functional Representations

DGX agent

arXiv:2509.24496v3 Announce Type: replace Abstract: The explosive growth of large language models (LLMs) has created a vast but opaque landscape: millions of models exist, yet their evolutionary relat

researcharxiv-cs-lg
4 May 2026
Safety

Online Self-Calibration Against Hallucination in Vision-Language Models

DGX agent

arXiv:2605.00323v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) often suffer from hallucinations, generating descriptions that include visual details absent from the input image.

safetyarxiv-cs-cv
4 May 2026
Model Releases

AutoSurfer -- Teaching Web Agents through Comprehensive Surfing, Learning, and Modeling

DGX agent

arXiv:2604.27253v1 Announce Type: new Abstract: Recent advances in multimodal large language models (LLMs) have revolutionized web agents that can automate complex tasks on websites. However, their ac

model-releasesarxiv-cs-ai
1 May 2026
Hardware

Benchmarking Deep Learning Models for Object Detection on Edge Computing Devices

DGX agent

arXiv:2409.16808v1 Announce Type: cross Abstract: Modern applications, such as autonomous vehicles, require deploying deep learning algorithms on resource-constrained edge devices for real-time image

hardwarearxiv-cs-lg
1 May 2026
Research

Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models

DGX agent

arXiv:2604.27251v1 Announce Type: cross Abstract: Large Language Models (LLMs) are known to acquire reasoning capabilities through shared inference patterns in pre-training data, which are further eli

researcharxiv-cs-ai
1 May 2026
Model Releases

DPN-LE: Dual Personality Neuron Localization and Editing for Large Language Models

DGX agent

arXiv:2604.27929v1 Announce Type: new Abstract: With the widespread adoption of large language models (LLMs), understanding their personality representation mechanisms has become critical. As a novel

model-releasesarxiv-cs-cl
1 May 2026
Safety

In-context Learning vs. Instruction Tuning: The Case of Small and Multilingual Language Models

DGX agent

arXiv:2503.01611v3 Announce Type: replace Abstract: Instruction following is a critical ability for Large Language Models to perform downstream tasks. The standard approach to instruction tuning has r

safetyarxiv-cs-cl
1 May 2026
← Previous
1…114115116117118…1030
Next →