AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

Economic Evaluations of Language Models

DGX agent

arXiv:2607.19375v1 Announce Type: cross Abstract: Language models perform economically valuable work, yet they are not currently assessed for how well they perform every economically valuable task. We

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

MissingBench-Verified: Probing Vision-Language Models' Inability to Detect Missing Object Parts

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.18673v2 Announce Type: new Abstract: Vision Language Models (VLMs) are well known for hallucinating non-existent objects in images. Objects with missing parts present a unique challenge for

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Post-Training in Time Series Foundation Models: A Unifying Framework

DGX agent

arXiv:2607.20002v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have emerged as general-purpose models for time series analysis, but pretraining alone is often insufficient for

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

ROMS-IMLE: A Minimalist Approach to Competitive Single-Step Generative Modelling

DGX agent

arXiv:2607.19332v1 Announce Type: cross Abstract: Generative models have undergone many generations of evolution, from VAEs/GANs to diffusion/flow matching. Along the way, the underlying techniques ha

model-releasesarxiv-cs-cv
23 Jul 2026
Research

The C-index illusion: discrimination without calibration in published survival models

DGX agent

arXiv:2607.19526v1 Announce Type: new Abstract: 'Stop Chasing the C-index when Evaluating Survival Analysis Models' (ICML 2026, Spotlight) argued normatively, on synthetic data, that evaluating surviv

researcharxiv-cs-lg
23 Jul 2026
Model Releases

DarwinLM: Evolutionary Structured Pruning of Large Language Models

DGX agent

arXiv:2502.07780v4 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved significant success across various NLP tasks. However, their massive computational costs limit their wide

model-releasesarxiv-cs-lg
16 Jul 2026
Agents

What Models Express, Suppress, and Resist: Auditing Open-Weight LLMs with Persona Vectors

DGX agent

arXiv:2607.13162v1 Announce Type: cross Abstract: What a language model will and will not do is largely set during post-training, but which behaviors it expresses, hides, or resists is not revealed by

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

ABot-N1: Toward a General Visual Language Navigation Foundation Model

DGX agent

arXiv:2607.10383v2 Announce Type: replace-cross Abstract: Visual Language Navigation foundation models aim to unify deep reasoning for grounded spatial decisions with broad versatility for diverse emb

model-releasesarxiv-cs-ai
15 Jul 2026
Research

Audio-Native Speech Recognition with a Frozen Discrete-Diffusion Language Model

DGX agent

arXiv:2607.13013v1 Announce Type: new Abstract: Automatic speech recognition is dominated by autoregressive decoders that emit one token at a time. We ask whether a discrete diffusion language model c

researcharxiv-cs-ai
15 Jul 2026
Model Releases

DiTailed: Ensuring Visual Object Consistency in Text-Image-to-Image Flow Matching Models

DGX agent

arXiv:2607.12539v1 Announce Type: new Abstract: Despite remarkable progress in text-guided image editing, generative models frequently fail to preserve visual object consistency, defined as the preser

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Evaluating Health Misinformation in Low-Resource Languages: Integrating Small Language Models with a Culturally-Sensitive Responsible NLP Framework (Bangla as a Case Study)

DGX agent

arXiv:2607.12336v1 Announce Type: cross Abstract: Artificial Intelligence (AI) technologies, while serving as a foundational enabler for modern social media and digital health services, exert a bivale

model-releasesarxiv-cs-ai
15 Jul 2026
Safety

Growing a Tail: Increasing Output Diversity in Large Language Models

DGX agent

arXiv:2411.02989v2 Announce Type: replace Abstract: How diverse are the outputs of large language models when diversity is desired? We examine the diversity of responses of several language models to

safetyarxiv-cs-cl
15 Jul 2026
Research

The Computational Basis of Confidence in Large Language Models

DGX agent

arXiv:2607.12447v1 Announce Type: cross Abstract: Reliable confidence -- the probability that a model's own answer is correct -- is essential for the trustworthy deployment of language models. Existin

researcharxiv-cs-ai
15 Jul 2026
Agents

The Emerging Paradigm of Geospatial Foundation Models: From Pre-Training to Agentic Reasoning

DGX agent

arXiv:2607.12177v1 Announce Type: new Abstract: The analysis of satellite and aerial imagery has entered a new era with the advent of foundation models. This paper describes the concept of Geospatial

agentsarxiv-cs-ai
15 Jul 2026
Local Ai

An exact information theory of generalization phase transitions in Bayesian diffusion models

DGX agent

arXiv:2607.08041v1 Announce Type: new Abstract: How diffusion models circumvent the curse of dimensionality to learn complex distributions over high dimensional spaces from a finite training set, inst

local-aiarxiv-cs-lg
10 Jul 2026
Model Releases

Beyond wheelchairs and blindfolds: Investigating disability stereotypes in T2I models with INCLUDE-BENCH

DGX agent

arXiv:2607.08515v1 Announce Type: new Abstract: Text-to-image (T2I) models have been shown to exhibit social biases. Prior work has mainly focused on gender, skin tone, and cultural representation wit

model-releasesarxiv-cs-cv
10 Jul 2026
Research

Robust Weighted Triangulation of Causal Effects Under Model Uncertainty

DGX agent

arXiv:2603.01119v2 Announce Type: replace-cross Abstract: A fundamental challenge in causal inference with observational data is correct specification of a causal model. When there is model uncertaint

researcharxiv-cs-ai
10 Jul 2026
Model Releases

Diffusion Models in Simulation-Based Inference: A Tutorial Review

DGX agent

arXiv:2512.20685v3 Announce Type: replace-cross Abstract: Diffusion models have recently emerged as powerful learners for simulation-based inference (SBI), enabling fast and accurate estimation of lat

model-releasesarxiv-cs-lg
9 Jul 2026
Safety

Gradient-Based Speech-to-Text Alignment for Any ASR Model: From CTC to Speech LLMs

DGX agent

arXiv:2607.06831v1 Announce Type: cross Abstract: Speech-to-text alignment means finding the temporal boundaries of each word in the audio. Some models provide such an alignment directly and others do

safetyarxiv-cs-ai
9 Jul 2026
Safety

Open-Ended Scenario Reasoning for Specialist Model Adaptation

DGX agent

arXiv:2607.06625v1 Announce Type: cross Abstract: Process industries have accumulated validated specialist models, yet sensor drift, feedstock variation, and regime switching cause these models to deg

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

FourTune: Towards Fully 4-Bit Efficient Post-Training for Diffusion Models

DGX agent

arXiv:2607.05711v1 Announce Type: cross Abstract: Diffusion models have become a dominant paradigm for high-quality generative modeling, while post-training is essential for adapting them to diverse d

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Harnessing Generative Image Models for Training-Free Primitive Shape Abstraction

DGX agent

arXiv:2607.05568v1 Announce Type: cross Abstract: Representing 3D shapes as compact sets of geometric primitives is fundamental to robotics, simulation, and scene understanding. Generative image model

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Omni-RRM: Advancing Omni Reward Modeling via Automatic Rubric-Grounded Preference Synthesis

DGX agent

arXiv:2602.00846v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) struggle with alignment due to the limitations of existing reward models (RMs), which are predominantly vis

model-releasesarxiv-cs-cl
8 Jul 2026
Research

Performance Optimization and Comparative Analysis of Generative AI Models on Advanced Accelerators

DGX agent

arXiv:2607.05400v1 Announce Type: cross Abstract: Generative AI models, such as Large Language Models (LLMs) and diffusion models, have demonstrated impressive performance across a wide range of tasks

researcharxiv-cs-lg
8 Jul 2026
Model Releases

The relationship between reasoning and performance in large language models--o3 (mini) thinks harder, not longer

DGX agent

arXiv:2502.15631v2 Announce Type: replace-cross Abstract: Large language models have demonstrated remarkable progress in mathematical reasoning, leveraging chain-of-thought and reinforcement learning.

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Brand-as-Memory: Vision-Language Models Encode Causal, Mechanistically Localizable Credibility Priors for News Sources

DGX agent

arXiv:2607.03365v1 Announce Type: cross Abstract: Vision-language models (VLMs) increasingly read news and web content as images, where the publisher's identity is visually present. We show that VLMs

model-releasesarxiv-cs-ai
7 Jul 2026
Local Ai

Can Model Merging Improve Aggregation in DiLoCo?

DGX agent

arXiv:2607.03011v1 Announce Type: cross Abstract: Model merging techniques, which aggregate independently finetuned models into one to combine their capabilities, have become a topic of significant in

local-aiarxiv-cs-ai
7 Jul 2026
Tutorials

CollabEval: Statistically Efficient Collaborative Model Evaluation via Matrix Completion

DGX agent

arXiv:2607.05046v1 Announce Type: new Abstract: Evaluating generative AI models is a routine, but resource-intensive, process that is conducted over and over again during the course of model developme

tutorialsarxiv-cs-lg
7 Jul 2026
Model Releases

Criterion-Conditional In-Context Learning: Evaluating Criterion-Shift Adaptation in Vision-Language Models

DGX agent

arXiv:2607.02575v1 Announce Type: cross Abstract: Vision-language models can perform new tasks without parameter updates through in-context learning (ICL), whose core mechanism is utilizing the suppor

model-releasesarxiv-cs-ai
7 Jul 2026
Applications

Cross-device Collaborative Test-time Adaptation with Zeroth-order Optimization and Model Merging

DGX agent

arXiv:2607.02988v1 Announce Type: new Abstract: Test-time adaptation (TTA) mitigates domain shifts by using incoming test data to update a model on the fly. The majority of TTA methods require resourc

applicationsarxiv-cs-cv
7 Jul 2026
Research

EvoXplain: When Machine Learning Models Agree on Predictions but Disagree on Why -- Measuring Mechanistic Multiplicity Across Training Runs

DGX agent

arXiv:2512.22240v5 Announce Type: replace-cross Abstract: Machine learning models are primarily judged by predictive performance, especially in applied genomics, where explanations are read as biologi

researcharxiv-cs-ai
7 Jul 2026
Model Releases

Multiplayer Interactive World Models with Representation Autoencoders

DGX agent

arXiv:2607.05352v1 Announce Type: cross Abstract: We introduce the first multiplayer world model for highly dynamic environments governed by complex physical interactions. Whereas single-player world

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

No Reliable Evidence of Self-Reported Sentience in Small Large Language Models

DGX agent

arXiv:2601.15334v2 Announce Type: replace-cross Abstract: Whether language models possess sentience has no empirical answer. But whether they believe themselves to be sentient can, in principle, be te

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Spectral Rewiring for Exploration, Purification, and Model Merging

DGX agent

arXiv:2607.03065v1 Announce Type: cross Abstract: Reinforcement learning has become a standard post-training recipe for large language models, but dense full-parameter updates create two deployment-re

model-releasesarxiv-cs-ai
7 Jul 2026
Local Ai

Trust-Region Noise Search for Black-Box Alignment of Diffusion and Flow Models

DGX agent

arXiv:2603.14504v2 Announce Type: replace-cross Abstract: Optimizing the noise samples of diffusion and flow models is an increasingly popular approach to align these models to target rewards at infer

local-aiarxiv-cs-ai
7 Jul 2026
Model Releases

Comparing Large Language Models on Scrum Certification-Style Questions: Accuracy, Stability, and Error Patterns

DGX agent

arXiv:2607.00048v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in exam- and certification-style question answering tasks, where their ability to retrieve, interpr

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Steal the Patch Size: Adversarially Manipulate Vision-Language Models

DGX agent

arXiv:2607.00174v1 Announce Type: new Abstract: We present a black-box model-stealing attack that recovers private vision-tokenizer configurations of deployed vision-language models (VLMs), including

model-releasesarxiv-cs-cv
2 Jul 2026
Research

Utilizing Earth Foundation Models to Enhance the Simulation Performance of Hydrological Models with AlphaEarth Embeddings

DGX agent

arXiv:2601.01558v2 Announce Type: replace-cross Abstract: Predicting river flow in places without streamflow records is challenging because basins respond differently to climate, terrain, vegetation,

researcharxiv-cs-ai
2 Jul 2026
Safety

How Should World Models Be Evaluated for Embodied Decision-Making? A Decision-Making-Centric Position

DGX agent

arXiv:2606.15032v2 Announce Type: replace Abstract: World models have become a central abstraction in modern AI. The term now refers to several different objects: action-conditioned environment models

safetyarxiv-cs-lg
30 Jun 2026
Model Releases

MACROCAST: A Vintage-Consistent Time Series Foundation Model for Real-Time Macroeconomic Forecasting

DGX agent

arXiv:2606.28670v1 Announce Type: cross Abstract: We introduce MACROCAST, a lightweight Time Series Foundation Model (TSFM) for real-time macroeconomic forecasting. Existing TSFMs suffer from data lea

model-releasesarxiv-cs-ai
30 Jun 2026
Research

On Test-Time Scaling for Vision-Language Models

DGX agent

arXiv:2606.28864v1 Announce Type: new Abstract: Test-time scaling is a paradigm where large models use additional compute at inference to achieve better performance, without changing model weights. Wh

researcharxiv-cs-cv
30 Jun 2026
Model Releases

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models

DGX agent

arXiv:2606.29196v1 Announce Type: cross Abstract: Do language models know when they are being tested? This question matters for AI safety: a model that recognises an evaluation context could alter its

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models

DGX agent

arXiv:2606.29815v1 Announce Type: new Abstract: Evaluating code large language models (Code LLMs) requires reliable detection of data leakage, where benchmark performance is artificially inflated by e

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

The Digital Afterlife of Empires: Four Language Models Converge on the Same Imperial Cartography of Writing

DGX agent

arXiv:2606.28325v1 Announce Type: cross Abstract: Large language models process the world's writing systems with radical inequality. We constructed the Digital Script Representation Index (DSRI), a se

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

When the Prompt Becomes Visual: Vision-Centric Jailbreak Attacks for Large Image Editing Models

DGX agent

arXiv:2602.10179v2 Announce Type: replace-cross Abstract: Recent advances in large image editing models have shifted the paradigm from text-driven instructions to vision-prompt editing, where user int

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models

DGX agent

arXiv:2606.26566v1 Announce Type: cross Abstract: Adversarial evaluation of AI systems has matured along four largely disconnected tracks: diffusion-based attacks on text and large language models (LL

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Beyond Perplexity: UTF-8 Validity in Byte-aware Language Models

DGX agent

arXiv:2606.14122v2 Announce Type: replace Abstract: Byte-level tokenization enables language models to handle any Unicode input, but models can generate invalid UTF-8 sequences when encountering rare

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

DualEval: Joint Model-Item Calibration for Unified LLM Evaluation

DGX agent

arXiv:2606.26429v1 Announce Type: cross Abstract: Current LLM evaluation relies on two complementary but often disconnected signals: static benchmarks with objective correctness labels and arena-style

model-releasesarxiv-cs-cl
26 Jun 2026
← Previous
1…2324252627…1021
Next →