AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
13 Aug 2026

Large Language Models Reproduce Racial Stereotypes When Used for Text Annotation

SafetyDGX agent

arXiv:2603.13891v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used for automated text annotation in tasks ranging from academic research to content moderation

Motion-as-Prompt: Enhancing Motion Reasoning in Multimodal Large Language Models via Motion-Guided Cross-Frame Visual Prompting

Model ReleasesDGX agent

arXiv:2608.11655v1 Announce Type: cross Abstract: Motion-centric video reasoning is fundamental to interactive applications such as robotic manipulation and autonomous navigation. However, multimodal

Ripple-Pivot Search: Active Parallel Decoding for Diffusion Large Language Models

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.11742v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive language models, offering the potential for substant

Social Meaning in Large Language Models: Structure, Magnitude, and Pragmatic Prompting

ApplicationsDGX agent

arXiv:2604.02512v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly exhibit human-like patterns of pragmatic and social reasoning. This paper addresses two related ques

12 Aug 2026

Agentic AI infrastructure shifts enterprise focus from model choice to platform control

AgentsDGX agent

As agentic AI infrastructure moves from experimentation into production, enterprises are confronting a more complex question than which model to use: how to control the cost, data exposure and infrast

An Exact Instrument for State Usage in Selective State-Space Models, and the Input-Driven Migration It Reveals

ResearchDGX agent

arXiv:2607.11796v2 Announce Type: replace Abstract: Selective state-space models such as Mamba route information through a bank of first-order modes whose input coupling is set by a learned selection

Cross-View Feature Matching: Survey, Benchmarking, and Foundation-Model Perspectives

ResearchDGX agent

arXiv:2608.11093v1 Announce Type: cross Abstract: Cross-view feature matching aims to establish reliable correspondences across images with large viewpoint variations. Over the past decade, the field

Decodable But Not Detachable: Training Data Granularity Determines Parametric Modularity in Large Language Models

ResearchDGX agent

arXiv:2608.10214v1 Announce Type: new Abstract: Do large language models contain domain-specific parametric shells: concentrated, causally necessary neuron populations whose removal selectively degrad

DEFT: Data-Efficient Frequency-domain Top-k Sampling via Inverse Discrete Fourier Transform for Spatiotemporal Dynamical Systems Modeling

ResearchDGX agent

arXiv:2608.11019v1 Announce Type: new Abstract: Modeling spatiotemporal dynamical systems governed by partial differential equations (PDEs) poses two major challenges: it either requires expensive phy

Energy and Performance Benchmarking of Deep Learning Models for Breast Cancer Detection

ResearchDGX agent

arXiv:2608.09996v1 Announce Type: cross Abstract: Recent advances in machine learning have greatly improved breast cancer detection, enabling more accurate and timely diagnosis. Deep learning (DL) mod

Evaluating Semantic and Spatial Guidance for Foundation Model Segmentation of Small-Scale PV in Remote Sensing Imagery

ResearchDGX agent

arXiv:2608.10801v1 Announce Type: new Abstract: Spatio-temporal PV data are essential for understanding adoption processes in off-grid regions, yet such data remain largely unavailable. Automated segm

Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence

HardwareDGX agent

arXiv:2608.10720v1 Announce Type: new Abstract: Omni-modal dialogue models can understand multimodal inputs and synthesize spoken replies, yet their responses remain visually disembodied. We introduce

Foundation Model-Enabled Efficient Data Sampling (FEEDS): A label-efficient training strategy for pan-cancer, multi-tracer PET/CT datasets

ResearchDGX agent

arXiv:2608.11076v1 Announce Type: new Abstract: Automated lesion segmentation in whole-body PET/CT imaging can assist clinicians with cancer detection, staging, and treatment planning across radiotrac

INSIDE the Student's Mind: Jointly Modeling Latent Reasoning and Action in LLM Student Simulators

SafetyDGX agent

arXiv:2608.10492v1 Announce Type: new Abstract: Large Language Model (LLM)-based simulators often reproduce observable actions but fail to capture the underlying reasoning behind them. In education, w

MazzikaAI: A knowledge-based performance-to-prompt compiler for real-time Arabic maqam accompaniment with a streaming text-to-music model

ApplicationsDGX agent

arXiv:2608.10360v1 Announce Type: cross Abstract: Arabic maqam music microtonal, modal, and built on ornamented call and response is among the traditions most underserved by generative music models, w

Physics-informed Diffusion Generative Model for Time-Series Data Synthesis in Dynamic Systems

ResearchDGX agent

arXiv:2608.10941v1 Announce Type: new Abstract: Industrial time-series signals, such as turbine temperature and rotational speed in aero-engines, are essential for monitoring the health and operationa

The Gaussian-Multinoulli Restricted Boltzmann Machine: A Potts Model Extension of the GRBM

Model ReleasesDGX agent

arXiv:2505.11635v2 Announce Type: cross Abstract: Many real-world tasks, from associative memory to symbolic reasoning, benefit from discrete, structured representations that standard continuous laten

Toward the Cognitive--Physical Limits of Embodied Intelligence through a World-Model-Centric Autonomous Racing Agent

Local AiDGX agent

arXiv:2608.10618v1 Announce Type: new Abstract: Embodied artificial intelligence aims to develop agents that perceive, reason, and act through continuous interaction with the physical world. However,

Towards Sustainable Artificial Intelligence: A Comprehensive Review and Comparative Analysis of Deep Learning Models' Carbon Footprint

ResearchDGX agent

arXiv:2608.09998v1 Announce Type: new Abstract: Artificial Intelligence (AI) and Machine Learning (ML) have become powerful tools for supporting and automating complex human tasks. Despite their benef

You chose the best model. Why is your agent still failing?

AgentsDGX agent

Public benchmarks can show how a model performs in general. Production reliability depends on the context and harness around it, which only your team can evaluate against its own data, workflows, and

11 Aug 2026

A foundation model of numerical intelligence with cross-disciplinary generalization

ResearchDGX agent

arXiv:2607.28432v2 Announce Type: replace Abstract: Intelligence is commonly understood as the ability to acquire and apply knowledge, adapt to unfamiliar situations and solve new problems. Large lang

An invertible generative model for forward and inverse problems

ResearchDGX agent

arXiv:2509.03910v2 Announce Type: replace-cross Abstract: We formulate inverse problems in a Bayesian framework and aim to train an invertible generative model that is capable of simulation (i.e., sam

Are Latent Reasoning Models Easily Interpretable?

ResearchDGX agent

arXiv:2604.04902v2 Announce Type: replace Abstract: Latent reasoning models (LRMs) have attracted significant research interest due to their low inference cost (relative to explicit reasoning models)

Diminishing Returns of Intelligence: The Non-Linear Relationship Between LLM Scale and User Perception in Short-Duration Open-Ended Social Human-Robot Interactions

Model ReleasesDGX agent

arXiv:2608.08320v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to drive embodied social agents, yet it remains unclear whether larger models improve user perception

DoGMA: A Central-Dogma-Guided Foundation Model for Multi-Omics Alignment and Multi-Task Learning in Oncology

SafetyDGX agent

arXiv:2608.08148v1 Announce Type: cross Abstract: Attention mechanisms have been widely utilized in modern deep learning, and many existing multi-omics models inherit their conventional use to allow u

Emotion in an active inference model of human driving

ResearchDGX agent

arXiv:2608.07480v1 Announce Type: new Abstract: Active inference has emerged as a principled framework for modeling adaptive behavior by balancing goal-directed action with uncertainty reduction. It h

Ethical Framework for Responsible Foundational Models in Medical Imaging

SafetyDGX agent

arXiv:2406.11868v2 Announce Type: replace-cross Abstract: The emergence of foundational models represents a paradigm shift in medical imaging, offering extraordinary capabilities in disease detection,

Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Threats and Safeguards

SafetyDGX agent

arXiv:2608.07535v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) integrate heterogeneous modalities through modality alignment and fusion, enabling stronger understanding an

HOPPER: Learnable Hop Extraction for Linearized Graph Sequence Models

Model ReleasesDGX agent

arXiv:2608.09031v1 Announce Type: new Abstract: Graph neural networks typically propagate information through repeated message-passing layers, coupling the distance over which information travels with

Latent World Models with Monotone Planning Costs for Image-Goal Navigation

ApplicationsDGX agent

arXiv:2608.09073v1 Announce Type: new Abstract: Image-goal navigation with latent world models requires not only accurate future prediction, but also a planning cost that reliably ranks candidate acti

Leveraging generative models to assist Monte Carlo sampling

TutorialsDGX agent

arXiv:2608.07648v1 Announce Type: cross Abstract: Sampling high-dimensional probability distributions is a central task in scientific computing, with applications ranging from Bayesian inference to st

Listen, See and Track: Spatio-Temporal Audio-Visual Sound Event Reasoning for Omni-Modal Language Models

Model ReleasesDGX agent

arXiv:2608.09435v1 Announce Type: new Abstract: Understanding dynamic sound sources requires jointly determining what produces a sound, where the source is located, and how it moves over time. Yet exi

LLMs Remember First, Forget Last: Dual-Process Interference in Large Language Models

SafetyDGX agent

arXiv:2603.00270v3 Announce Type: replace-cross Abstract: Large language models can process millions of tokens, yet how they handle conflicting information within context remains poorly understood. Fr

LookME: Lookup-Based Multimodal Embeddings for Layer Injection in Vision-Language Models

ResearchDGX agent

arXiv:2607.16305v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have achieved strong progress in multimodal understanding. However, scaling dense or sparse Mixture-of-Experts (

MaxModShift: Model Privacy via Designed Shifts

TutorialsDGX agent

arXiv:2608.09328v1 Announce Type: new Abstract: Model learning by an eavesdropper is treated as an estimation problem in a federated environment. The Fisher Information Matrix for the eavesdropper's e

Multilingual Agent-Based World Modeling for Social Science

Model ReleasesDGX agent

arXiv:2512.07195v2 Announce Type: replace-cross Abstract: Multi-agent role-playing has recently shown promise for studying social behavior with language agents, but existing simulations are mostly mon

One Model to Magnify Them All: Efficient Scale-Invariant Histopathology via Conditional Normalization and Continuous Magnification Training

ResearchDGX agent

arXiv:2608.09403v1 Announce Type: new Abstract: Whole slide images (WSIs) in digital histopathology are acquired at discrete magnification levels encoding complementary diagnostic information from glo

Performance of large language models in the optical diagnosis of colorectal polyps

Model ReleasesDGX agent

arXiv:2608.07543v1 Announce Type: cross Abstract: Background and Study Aims: Accurate optical diagnosis of colorectal polyps guides resection strategy and surveillance, with multimodal large language

ReMIND: Orchestrating Modular Large Language Models for Controllable Serendipity A REM-Inspired System Design for Emergent Creative Ideation

ResearchDGX agent

arXiv:2601.07121v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used not only for problem solving but also for creative ideation; however, generating ideas that

Renormalising Generative Models for Active Inference: Foundations, Derivations, and Verification

ResearchDGX agent

arXiv:2608.09512v1 Announce Type: new Abstract: Active inference offers a unified framework for perception, learning, and action, but scaling discrete active-inference models to rich spatial and tempo

Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard

HardwareDGX agent

NVIDIA NeMo Switchyard is a routing platform that directs AI agent workloads to the most suitable specialized or frontier model for each step of a task, balancing performance, cost, and latency. It of

Source: Trajectory, founded by ex-DeepMind, Apple, OpenAI, and Meta staffers to build continual learning models, raised 40M led by Sequoia at a 300M valuation (Stephanie Palazzolo/The Information)

IndustryDGX agent

Stephanie Palazzolo / The Information: Source: Trajectory, founded by ex-DeepMind, Apple, OpenAI, and Meta staffers to build continual learning models, raised 40M led by Sequoia at a 300M valuation —

Toward Mask Annotation-Free Surgical Instrument Segmentation from Endoscopic Images Using Text-Prompted Segment Anything Model 3 (SAM3)

Model ReleasesDGX agent

arXiv:2608.08844v1 Announce Type: new Abstract: Surgical instrument segmentation is a fundamental task for computer-assisted interventions, yet most existing methods rely on pixel-level annotations or

10 Aug 2026

CellWorld: From Gene-Level Reconstruction to Latent Cell Prediction in Spatial Transcriptomics Foundation Models

ResearchDGX agent

arXiv:2608.06659v1 Announce Type: new Abstract: This paper shows that latent-space predictive pretraining can provide a scalable route to foundation models for spatial transcriptomics. Existing spatia

Convergence of Diffusion Models Under the Manifold Hypothesis in High-Dimensions

ResearchDGX agent

arXiv:2409.18804v3 Announce Type: replace-cross Abstract: Denoising Diffusion Probabilistic Models (DDPM) are powerful state-of-the-art methods used to generate synthetic data from high-dimensional da

Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model

ResearchDGX agent

arXiv:2608.07361v1 Announce Type: cross Abstract: Vision-language-action (VLA) models route driving decisions through a deep language model, but it is unclear how much of that depth the action itself

Faster Query-Key Learning Sharpens Attention in Self-Attention Models

Model ReleasesDGX agent

arXiv:2608.06776v1 Announce Type: new Abstract: A standard self-attention layer consists of two interacting circuits: the query-key circuit that governs attention allocation, and the output-value circ

How Molecular Generative Models Organize Molecular Identity

ResearchDGX agent

arXiv:2608.06956v1 Announce Type: new Abstract: Generative models for matter are often evaluated as samplers over output representations, and their latent spaces are commonly used as proxies for navig

Native Long Video Understanding Models locally?

Model ReleasesDGX agent

I've been building a personal project and wanted to check with the community on multi-modal inputs since I can't find a lot of material around this online. Ultimately I'm trying to build something tha

Separating Decision-Rule Misalignment from Readout-Coverage Limitations in Speech Language Models

ResearchDGX agent

arXiv:2608.06409v1 Announce Type: cross Abstract: Speech language models are increasingly evaluated on paralinguistic tasks by the accuracy of prompted answers, but answer accuracy combines failures a

9 Aug 2026

In-depth look at OpenAI's model training, dangerous decisions, and cluelessness before the HuggingFace hack; despite delaying Astra, OpenAI still doesn't get it (Zvi Mowshowitz/Don't Worry About the Vase)

IndustryDGX agent

Zvi Mowshowitz / Don't Worry About the Vase: In-depth look at OpenAI's model training, dangerous decisions, and cluelessness before the HuggingFace hack; despite delaying Astra, OpenAI still doesn't g

是我见过OCR效果最好的一个了,人眼都难以分辨的,它能处理的十分准确,而且速度非常的快。llamaindex 真不愧是文档解析界的一哥啊。

Model ReleasesDGX agent

是我见过OCR效果最好的一个了,人眼都难以分辨的,它能处理的十分准确,而且速度非常的快。llamaindex 真不愧是文档解析界的一哥啊。 The best 'raw' frontier model for document parsing is gemini 3 flash, but the issue is that since then the flash models have gotte

8 Aug 2026

Now we have a timeline of the OpenAI accidental attack against Hugging Face

Model ReleasesDGX agent

My comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi

7 Aug 2026

AI Playing Business Games: Benchmarking Large Language Models on Managerial Decision-Making in Dynamic Simulations

Model ReleasesDGX agent

arXiv:2509.26331v2 Announce Type: replace Abstract: The rapid advancement of LLMs sparked significant interest in their potential to augment or automate managerial functions. One of the most recent tr

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs

Model ReleasesDGX agent

arXiv:2607.18056v2 Announce Type: replace Abstract: Frontier large language models (LLMs) are increasingly integrated into scientific workflows, yet their growing biological capabilities may outpace c

Clinician input steers AI toward accurate and harmful recommendations

Model ReleasesDGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

SafetyDGX agent

arXiv:2608.05695v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly invoke external tools and interact with real-world systems, unsafe actions may cause irreversible cons

LAWM-3D: Learning 3D-Aware Latent Actions from Human Videos for Generalizable Robot World Models

SafetyDGX agent

arXiv:2608.05706v1 Announce Type: new Abstract: World models enable agents to perform forward rollout and planning without real-world interaction. However, their application in open-world embodied int

On the Anisotropy of Score-Based Generative Models

ResearchDGX agent

arXiv:2510.22899v2 Announce Type: replace Abstract: We investigate the role of network architecture in shaping the inductive biases of modern score-based generative models. To this end, we introduce t

The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping

Model ReleasesDGX agent

arXiv:2608.06361v1 Announce Type: new Abstract: Real-world video benchmarks provide broad coverage, but their fixed clips entangle event count, rate, duration, and visual complexity, making failure mo

← Previous
1…123124125126127…1009
Next →