AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,903 results
Research

Behavioral Geometric Supervision Aligns Video Foundation Models with Human Social Perception

DGX agent

arXiv:2510.01502v2 Announce Type: replace-cross Abstract: Current video foundation models, including the strongest self-supervised models such as V-JEPA2, fail to capture how humans organize social in

researcharxiv-cs-cv
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

DGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

No One Knows the State of the Art in Geospatial Foundation Models

DGX agent

arXiv:2605.12678v1 Announce Type: new Abstract: Geospatial foundation models (GFMs) have been proposed as generalizable backbones for disaster response, land-cover mapping, food-security monitoring, a

model-releasesarxiv-cs-cv
14 May 2026
Research

Sample-Efficient Optimisation over the Outputs of Generative Models

DGX agent

arXiv:2509.23800v3 Announce Type: replace-cross Abstract: Modern generative AI models, such as diffusion and flow matching models, can sample from rich data distributions. However, many applications,

researcharxiv-cs-lg
14 May 2026
Applications

Bayesian Surrogate Training on Multiple Data Sources: A Hybrid Modeling Strategy

DGX agent

arXiv:2412.11875v3 Announce Type: replace-cross Abstract: Surrogate models are often used as computationally efficient approximations to complex simulation models, enabling tasks such as solving inver

applicationsarxiv-cs-lg
13 May 2026
Model Releases

Large-Small Model Collaboration for Farmland Semantic Change Detection

DGX agent

arXiv:2605.12282v1 Announce Type: new Abstract: Farmland Semantic Change Detection (SCD) is essential for cultivated land protection, yet existing benchmarks and models remain insufficient for fine-gr

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution

DGX agent

arXiv:2504.14129v4 Announce Type: replace Abstract: The challenge of tracing the source attribution of forged faces has gained significant attention due to the rapid advancement of generative models.

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

Strategically Deceptive Model Deployment in Performative Prediction

DGX agent

arXiv:2506.09044v2 Announce Type: replace Abstract: Machine Learning systems are increasingly deployed in decision-making settings that shape user behavior and, in turn, the data on which future decis

tutorialsarxiv-cs-lg
13 May 2026
Research

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

DGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

researcharxiv-cs-cl
13 May 2026
Research

Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing

DGX agent

arXiv:2605.10794v1 Announce Type: cross Abstract: Language models are deployed in settings that require compartmentalization: system prompts should not be disclosed, chain-of-thought reasoning is hidd

researcharxiv-cs-ai
12 May 2026
Model Releases

Language Models Without a Trainable Input Embedding Table: Learning from Fixed Minimal Binary Token Codes

DGX agent

arXiv:2605.09751v1 Announce Type: new Abstract: Trainable input embedding tables are a standard component of modern language models. We ask whether they are actually necessary at the input interface.

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Layer Collapse in Diffusion Language Models

DGX agent

arXiv:2605.06366v2 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as competitive alternatives to autoregressive (AR) language models, yet differences in their

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Learning Multi-Indicator Weights for Data Selection: A Joint Task-Model Adaptation Framework with Efficient Proxies

DGX agent

arXiv:2605.09665v1 Announce Type: cross Abstract: Data selection is a key component of efficient instruction tuning for large language models, as recent work has shown that data quality often matters

model-releasesarxiv-cs-ai
12 May 2026
Safety

Machine Unlearning on Pre-trained Models by Residual Feature Alignment Using LoRA

DGX agent

arXiv:2411.08443v2 Announce Type: replace-cross Abstract: Machine unlearning is an emerging technology that removes a subset of the training data from a trained model without significantly affecting t

safetyarxiv-cs-cv
12 May 2026
Model Releases

Model Capacity Determines Grokking through Competing Memorisation and Generalisation Speeds

DGX agent

arXiv:2605.09724v1 Announce Type: new Abstract: Existing accounts of grokking explain the phenomena in terms of mechanistic frameworks such as circuit efficiency or lazy-to-rich transitions. However,

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Models got an order of magnitude better at following instructions in one year

DGX agent

A year ago, frontier models started losing track of instructions somewhere around 200–300 simultaneous constraints. With 2026 models, that ceiling is closer to 2,000 — an order-of-magnitude jump. We r

model-releasesarize-ai
12 May 2026
Model Releases

Preserve and Personalize: Personalized Text-to-Image Diffusion Models without Distributional Drift

DGX agent

arXiv:2505.19519v3 Announce Type: replace Abstract: Personalizing text-to-image diffusion models involves integrating novel visual concepts from a small set of reference images while retaining the mod

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari

DGX agent

arXiv:2605.08578v1 Announce Type: cross Abstract: Developing generalist systems that retain human-like data efficiency is a central challenge. While world models (WMs) offer a promising path, existing

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Recursive Language Models

DGX agent

arXiv:2512.24601v3 Announce Type: replace Abstract: We study allowing large language models (LLMs) to process arbitrarily long prompts through the lens of inference-time scaling. We propose Recursive

model-releasesarxiv-cs-ai
12 May 2026
Safety

TripleWin: Fixed-Point Equilibrium Pricing for Data-Model Coupled Markets

DGX agent

arXiv:2511.03368v2 Announce Type: replace Abstract: The rise of the machine learning (ML) model economy has intertwined markets for training datasets and pre-trained models. However, most pricing appr

safetyarxiv-cs-lg
12 May 2026
Model Releases

Adapting Vision-Language Models for Neutrino Event Classification in High-Energy Physics

DGX agent

arXiv:2509.08461v3 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated their remarkable capacity to process and reason over structured and unstruct

model-releasesarxiv-cs-ai
11 May 2026
Agents

AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites

DGX agent

arXiv:2605.06841v1 Announce Type: new Abstract: In model-based learning, the agent learns behaviors by simulating trajectories based on world model predictions. Standard world models typically learn a

agentsarxiv-cs-ai
11 May 2026
Model Releases

Detecting Distillation Data from Reasoning Models

DGX agent

arXiv:2510.04850v3 Announce Type: replace-cross Abstract: Reasoning distillation has emerged as a prevailing paradigm for transferring reasoning capabilities from large reasoning models to small langu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Evaluating Large Language Models in Scientific Discovery

DGX agent

arXiv:2512.15567v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet prevailing science benchmarks probe decontextualized knowledge and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

How Big Should a Wireless Foundation Model Be?

DGX agent

arXiv:2605.07266v1 Announce Type: cross Abstract: Wireless foundation models are rapidly emerging as a key enabler of AI-native communication systems, yet a fundamental question remains unanswered: ho

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models

DGX agent

arXiv:2605.06672v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning and reasoning-tuned models such as DeepSeek-R1 are commonly assumed to reduce shallow heuristic biases by thinking care

model-releasesarxiv-cs-ai
11 May 2026
Safety

Post-training makes large language models less human-like

DGX agent

arXiv:2605.07632v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavi

safetyarxiv-cs-ai
11 May 2026
Model Releases

Query-efficient model evaluation using cached responses

DGX agent

arXiv:2605.07096v1 Announce Type: cross Abstract: Evaluating a new model on an existing benchmark is often necessary to understand its behavior before deployment. For modern evaluation frameworks, gen

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

THINKSAFE: Self-Generated Safety Alignment for Reasoning Models

DGX agent

arXiv:2601.23143v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve remarkable performance by leveraging reinforcement learning (RL) on reasoning tasks to generate long chain-of-

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Emergent Hierarchical Structure in Large Language Models: An Information-Theoretic Framework for Multi-Scale Representation

DGX agent

arXiv:2505.18244v3 Announce Type: replace Abstract: Why do language models from different architecture families respond so differently to the same perturbation? We argue that the answer is not scale,

model-releasesarxiv-cs-cl
7 May 2026
Safety

Enhancing the interpretability of spatially variable N2O model predictions with soft sensors during wastewater treatment

DGX agent

arXiv:2605.04082v1 Announce Type: new Abstract: Model-based solutions for nitrous oxide (N2O) emissions from wastewater treatment plants (WWTP) are informed by operational datasets designed to control

safetyarxiv-cs-lg
7 May 2026
Research

Gyan: An Explainable Neuro-Symbolic Language Model

DGX agent

arXiv:2605.04759v1 Announce Type: new Abstract: Transformer based pre-trained large language models have become ubiquitous. There is increasing evidence to suggest that even with large scale pre-train

researcharxiv-cs-cl
7 May 2026
Model Releases

Low-Rank Adaptation of Geospatial Foundation Models for Wildfire Mapping Using Sentinel-2 Data

DGX agent

arXiv:2605.04989v1 Announce Type: new Abstract: Wildfire burned-area mapping is essential for damage assessment, emissions modeling, and understanding fire-climate interactions across diverse ecologic

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Open-Source Image Editing Models Are Zero-Shot Vision Learners

DGX agent

arXiv:2605.04566v1 Announce Type: cross Abstract: Recent studies have shown that large generative models can solve vision tasks they were not explicitly trained for. However, existing evidence relies

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Probing Structural Mathematical Reasoning in Language Models with Algebraic Trapdoors

DGX agent

arXiv:2605.04352v1 Announce Type: new Abstract: We introduce a benchmark suite for evaluating structural mathematical reasoning in language models, built on subgroup-construction problems in SL(3, Z)

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Scalable Object Detection in the Car Interior With Vision Foundation Models

DGX agent

arXiv:2508.19651v2 Announce Type: replace Abstract: AI tasks in the car interior like identifying and localizing externally introduced objects is crucial for response quality of personal assistants. H

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?

DGX agent

arXiv:2605.03195v1 Announce Type: new Abstract: Modern coding agents increasingly delegate specialized subtasks to subagents, which are smaller, focused agentic loops that handle narrow responsibiliti

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

A Benchmark for Interactive World Models with a Unified Action Generation Framework

DGX agent

arXiv:2605.03941v1 Announce Type: new Abstract: Achieving Artificial General Intelligence (AGI) requires agents that learn and interact adaptively, with interactive world models providing scalable env

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving

DGX agent

arXiv:2509.17677v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance on mathematical reasoning under well-defined conditions. However, real-world engineering

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Expanding functional protein sequence space using high entropy generative models

DGX agent

arXiv:2605.03578v1 Announce Type: cross Abstract: Boltzmann Machines trained on evolutionary sequence data have emerged as a powerful paradigm for the data-driven design of artificial proteins. Howeve

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Seeking Information with RAG-Assistants: Does Model Size Matter in Human-AI Collaborations?

DGX agent

arXiv:2605.00964v1 Announce Type: cross Abstract: Much research on LLMs has focused on increasing benchmark performance. However, the evaluation of such models in real-world collaborative human-AI wor

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Agentopic: A Generative AI Agent Workflow for Explainable Topic Modeling

DGX agent

arXiv:2605.00833v1 Announce Type: new Abstract: Agentopic is a novel agent-based workflow for explainable topic modeling that leverages the reasoning capabilities of Large Language Models (LLMs). Exis

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE

DGX agent

arXiv:2605.02641v1 Announce Type: new Abstract: We present Mamoda2.5, a unified AR-Diffusion framework that seamlessly integrates multimodal understanding and generation within a single architecture.

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Statistically-Lossless Quantization of Large Language Models

DGX agent

arXiv:2605.02404v1 Announce Type: new Abstract: Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

TokenTiming: A Dynamic Alignment Method for Universal Speculative Decoding Model Pairs

DGX agent

arXiv:2510.15545v4 Announce Type: replace Abstract: Accelerating the inference of large language models (LLMs) has been a critical challenge in generative AI. Speculative decoding (SD) substantially i

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

WildTableBench: Benchmarking Multimodal Foundation Models on Table Understanding In the Wild

DGX agent

arXiv:2605.01018v1 Announce Type: new Abstract: Using multimodal foundation models to analyze table images is a high-value yet challenging application in consumer and enterprise scenarios. Despite its

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harne…

DGX agent

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harness that's truly model-agnostic, without compromising perform

model-releasesharrison-chase--x
4 May 2026
Research

Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision

DGX agent

arXiv:2605.00644v1 Announce Type: new Abstract: Energy-based models (EBMs) are a flexible class of deep generative models and are well-suited to capture complex dependencies in multimodal data. Howeve

researcharxiv-cs-lg
4 May 2026
← Previous
1…3839404142…1248
Next →