AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,428 results
Model Releases

Large Language Model-Powered Query-Driven Event Timeline Summarization in Industrial Search

DGX agent

arXiv:2605.27066v1 Announce Type: new Abstract: Understanding how events evolve over time is essential for search engines handling queries about trending news. We present QDET (Query-Driven Event Time

model-releasesarxiv-cs-cl
27 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Learning When to Think While Listening in Large Audio-Language Models

DGX agent

arXiv:2605.27190v1 Announce Type: cross Abstract: Recent advances in Large Audio-Language Models (LALMs) have made real-time, streaming spoken interaction increasingly practical. In this setting, reas

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LiveK12Bench: Have Large Multimodal Models Truly Conquered High School-level Examinations?

DGX agent

arXiv:2605.26781v1 Announce Type: new Abstract: Advanced Large Multimodal Models (LMMs) have demonstrated impressive performance in K-12 reasoning tasks, exhibiting great promise as intelligent tutors

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models

DGX agent

arXiv:2605.26895v1 Announce Type: cross Abstract: Normalization layers in modern large language models (LLMs) consist of a deterministic normalization operation and a learnable scale vector. While the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Reasoning, Code, or Both? How Large Language Models Handle Variations in Math Questions

DGX agent

arXiv:2605.26414v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve impressive accuracy on mathematical reasoning benchmarks, yet their performance drops when problems are modified wi

model-releasesarxiv-cs-ai
27 May 2026
Research

A Closer Look on Memorization in Tabular Diffusion Model: A Data-Centric Perspective

DGX agent

arXiv:2505.22322v3 Announce Type: replace Abstract: Diffusion models have shown strong performance in generating high-quality tabular data, but they carry privacy risks by reproducing exact training s

researcharxiv-cs-lg
26 May 2026
Safety

A comparative study of accuracy and rollout stability of temporal surrogate models

DGX agent

arXiv:2605.24868v1 Announce Type: new Abstract: Temporal surrogate models are effective for predicting chaotic dynamical systems where computational cost can be prohibitive. Several deep neural networ

safetyarxiv-cs-lg
26 May 2026
Research

Beyond Control-Flow: Integrating the Resource Perspective into Multi-Collaborative Process Modeling from Text

DGX agent

arXiv:2605.24546v1 Announce Type: new Abstract: Process modeling is a sub-domain of Business Process Management (BPM) focused on the translation of process artifacts into formal models. This task trad

researcharxiv-cs-ai
26 May 2026
Model Releases

Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models

DGX agent

arXiv:2605.25765v1 Announce Type: cross Abstract: Concept unlearning aims to erase a target concept from a pretrained text-to-image diffusion model without retraining. Closed-form methods are attracti

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CyberMaskQA: A Privacy-Aware Benchmark for Evaluating Large Language Models in Cybersecurity Question Answering

DGX agent

arXiv:2605.24765v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly applied to cybersecurity question answering (QA) for critical tasks such as incident response and vulner

model-releasesarxiv-cs-lg
26 May 2026
Safety

Do Understanding and Generation Fight? A Diagnostic Study of DPO for Unified Multimodal Models

DGX agent

arXiv:2603.17044v2 Announce Type: replace-cross Abstract: Unified multimodal models share a language model backbone for both understanding and generating images. Can DPO align both capabilities simult

safetyarxiv-cs-ai
26 May 2026
Agents

Drift-Resistant Navigation World Model with Anchored Epipolar Guidance

DGX agent

arXiv:2605.24761v1 Announce Type: cross Abstract: We propose Drift-Resistant Navigation World Model, a generative model that mitigates both perceptual drift and geometric drift in conventional rollout

agentsarxiv-cs-ro
26 May 2026
Model Releases

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model

DGX agent

arXiv:2510.10921v3 Announce Type: replace-cross Abstract: Fine-grained vision-language understanding requires precise alignment between visual content and linguistic descriptions, a capability that re

model-releasesarxiv-cs-ai
26 May 2026
Safety

Generative Visual Code Mobile World Models

DGX agent

arXiv:2602.01576v2 Announce Type: replace-cross Abstract: Mobile Graphical User Interface (GUI) World Models (WMs) offer a promising path for improving mobile GUI agent performance at train- and infer

safetyarxiv-cs-ai
26 May 2026
Safety

MARS: Margin and Semantic-Aware Data Augmentation for Reward Modeling

DGX agent

arXiv:2602.17658v2 Announce Type: replace-cross Abstract: Reward modeling is central to alignment pipelines such as RLHF, RLAIF, and PPO-based policy optimization, yet its reliability is constrained b

safetyarxiv-cs-ai
26 May 2026
Model Releases

Mix-MoE: Improving Multilingual Machine Translation of Large Language Models through Mixed MoEs

DGX agent

arXiv:2605.24681v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown great promise in multilingual machine translation (MT), even with limited bilingual supervision. However, fine

model-releasesarxiv-cs-ai
26 May 2026
Safety

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning

DGX agent

arXiv:2605.25210v1 Announce Type: cross Abstract: Diffusion models are increasingly used as powerful conditional generators, yet real deployments often involve multiple target distributions arising fr

safetyarxiv-cs-ai
26 May 2026
Model Releases

ScaleAcross Explorer: Exploring Communication Optimization for Scale-Across AI Model Training

DGX agent

arXiv:2605.24326v1 Announce Type: cross Abstract: The rapid scaling of large language model training requires distributing GPU resources across multiple data center buildings and regions. We refer to

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models

DGX agent

arXiv:2605.25510v1 Announce Type: new Abstract: Children increasingly have access to Large Language Models (LLMs), which may expose them to responses that are developmentally inappropriate or require

model-releasesarxiv-cs-cl
26 May 2026
Research

Transformer-based few-shot learning for modeling Electricity Consumption Profiles with minimal data across thousands of domains

DGX agent

arXiv:2408.08399v3 Announce Type: replace Abstract: Electricity Consumption Profiles (ECPs) are crucial for operating and planning power distribution systems, especially with the increasing number of

researcharxiv-cs-lg
26 May 2026
Model Releases

ViroBench: Benchmarking Nucleotide Foundation Models on Viral Genomics Tasks

DGX agent

arXiv:2605.25388v1 Announce Type: new Abstract: Nucleotide sequences constitute the fundamental genetic basis of biological systems, rendering viral genomic analysis critical for biomedical advancemen

model-releasesarxiv-cs-lg
26 May 2026
Research

Beyond Log Likelihood: Probability-Based Objectives for Supervised Fine-Tuning across the Model Capability Continuum

DGX agent

arXiv:2510.00526v3 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is the standard approach for post-training large language models (LLMs), yet it often shows limited generalization. We

researcharxiv-cs-cl
25 May 2026
Model Releases

Can AI Guess What You Know? Performance Comparison of Large Language Models for Human Domain Knowledge Estimation From Communication Logs

DGX agent

arXiv:2605.22971v1 Announce Type: new Abstract: Employees often struggle to identify ``who knows what,'' leading to organizational productivity losses. We investigate whether Large Language Models (LL

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models

DGX agent

arXiv:2605.23699v1 Announce Type: new Abstract: Video prediction is increasingly viewed as a path toward generalizable world models, yet it remains unclear whether these systems learn underlying causa

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Decomposition-Based Modular Conformal Prediction for Two-Stage Modeling

DGX agent

arXiv:2510.04406v2 Announce Type: replace-cross Abstract: Conformal prediction offers finite-sample coverage guarantees under minimal assumptions. However, existing methods treat the entire modeling p

model-releasesarxiv-cs-lg
25 May 2026
Research

DiLaDiff: Distilled Latent-Augmented Diffusion for Language Modeling

DGX agent

arXiv:2605.23605v1 Announce Type: cross Abstract: Diffusion language models intrinsically fail to capture correlations between decoded tokens, which leads to a harsh trade-off between sampling quality

researcharxiv-cs-ai
25 May 2026
Model Releases

How Far Are We from Generating Missing Modalities with Foundation Models?

DGX agent

arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Joint Model Parameter Scaling and Universal-Domain Data Integration for E-commerce Search Ranking

DGX agent

arXiv:2603.24226v3 Announce Type: replace-cross Abstract: Scaling studies for industrial search, advertising, and recommendation have largely emphasized enlarging model capacity or refining architectu

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Multi-SpatialMLLM: Multi-Frame Spatial Understanding with Multi-Modal Large Language Models

DGX agent

arXiv:2505.17015v2 Announce Type: replace-cross Abstract: Multi-modal large language models (MLLMs) have rapidly advanced in visual tasks, yet their spatial understanding remains limited to single ima

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Open Multimodal Datasets and Open-Source Software for Data-Driven Modeling of Multiphase Transport and Thermal Systems

DGX agent

arXiv:2605.23037v1 Announce Type: new Abstract: Data-driven modeling is becoming central to multiphase transport, electronics cooling, acoustic diagnostics, and thermal-fluid digital twins, but progre

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs

DGX agent

arXiv:2605.23157v1 Announce Type: new Abstract: The attack surface of a multimodal large language model (MLLM) is language-dependent in ways that reveal the mechanistic structure of alignment failures

model-releasesarxiv-cs-cl
25 May 2026
Research

Understanding Task Aggregation for Generalizable Ultrasound Foundation Models

DGX agent

arXiv:2603.18123v3 Announce Type: replace-cross Abstract: Foundation models promise to unify multiple clinical tasks within a single framework, but recent ultrasound studies report that unified models

researcharxiv-cs-ai
25 May 2026
Model Releases

Unextractable Protocol Models: Collaborative Training and Inference without Weight Materialization

DGX agent

arXiv:2605.23464v1 Announce Type: new Abstract: We consider a decentralized setup in which the participants collaboratively train and serve a large neural network, and where each participant only proc

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

VDE: Training-Free Accelerating Rectified Flow Model via Velocity Decomposition and Estimation

DGX agent

arXiv:2605.23381v1 Announce Type: new Abstract: Though rectified flow models have achieved remarkable performance in image, video, and 3D generation, their practical deployments are challenged by slow

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

When Symptoms Are Not Enough: Evidence-Weighting Patterns in Large Language Model Psychiatric Screening

DGX agent

arXiv:2605.23148v1 Announce Type: new Abstract: As demand for mental health care outpaces clinician-delivered assessment, scalable screening tools are increasingly needed. Large language models (LLMs)

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Billion-Scale Graph Foundation Models

DGX agent

arXiv:2602.04768v2 Announce Type: replace Abstract: Graph-structured data underpins many critical applications. While foundation models have transformed language and vision via large-scale pretraining

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

ChronoMedicalWorld: A Medical World Model for Learning Patient Trajectories from Longitudinal Care Data

DGX agent

arXiv:2605.21963v1 Announce Type: new Abstract: Long-horizon clinical simulation -- predicting how a patient's physiology evolves over years under specified interventions -- is central to chronic-dise

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

ChronoVAE-HOPE: Beyond Attention -- A Next-Generation VAE Foundation Model for Specialized Time Series Classification

DGX agent

arXiv:2605.22684v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have become a new component of the state-of-the-art in general time series forecasting. However, adapting them to

model-releasesarxiv-cs-lg
23 May 2026
Hardware

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

DGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

hardwarearxiv-cs-lg
23 May 2026
Model Releases

Tabular foundation models for robust calibration of near-infrared chemical sensing data

DGX agent

arXiv:2605.21544v1 Announce Type: new Abstract: Near-infrared spectroscopy is increasingly used as a rapid, non-destructive chemical sensing technology for the analysis of food, pharmaceutical, biolog

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answering

DGX agent

arXiv:2605.22099v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for grounding large language model (LLM) outputs in retrieved evidence, thereby

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark

DGX agent

arXiv:1709.03806v2 Announce Type: replace Abstract: Modern vision models have achieved strong object-recognition performance, yet it remains unclear whether their representations encode object-level s

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

I have to eat crow on this, in light of further information. whatever OpenAI spent on Erdos using a new model, apparently you can get GPT 5.…

DGX agent

I have to eat crow on this, in light of further information. whatever OpenAI spent on Erdos using a new model, apparently you can get GPT 5.5 to do something similar; @emollick’s presumably estimates

model-releasesgary-marcus--x
22 May 2026
Model Releases

Linear Dynamics in the RLVR Training of Large Language Models

DGX agent

arXiv:2601.04537v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven significant performance gains in reasoning-oriented large language models (LL

model-releasesarxiv-cs-cl
22 May 2026
Research

One prompt is not enough: Instruction Sensitivity Undermines Embedding Model Evaluation

DGX agent

arXiv:2605.22544v1 Announce Type: new Abstract: Instruction embedding models have become common among state-of-the-art models, however are evaluated using a single prompt per task. The single-point ev

researcharxiv-cs-cl
22 May 2026
Research

Probabilistic Attribution For Large Language Models

DGX agent

arXiv:2605.21726v1 Announce Type: new Abstract: The generative nature of Large Language Models (LLMs) is reflected in the conditional probabilities they compute to sample each response token given the

researcharxiv-cs-cl
22 May 2026
Model Releases

Reflective Prompt Tuning through Language Model Function-Calling

DGX agent

arXiv:2605.21781v1 Announce Type: new Abstract: Large language models (LLMs) have become increasingly capable of following instructions and complex reasoning, making prompting a flexible interface for

model-releasesarxiv-cs-cl
22 May 2026
Research

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation

DGX agent

arXiv:2605.21800v1 Announce Type: cross Abstract: World models are central to building agents that can reason, plan, and generalize beyond their training data. However, research on world models is cur

researcharxiv-cs-ro
22 May 2026
← Previous
1…8485868788…1259
Next →