AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

Noise2Params: Unification and Parameter Determination from Noise via a Probabilistic Event Camera Model

DGX agent

arXiv:2605.16317v1 Announce Type: new Abstract: Accurate, unified models for event cameras (ECs) remain elusive, hampering calibration and algorithm design. We develop a foundational probabilistic mod

model-releasesarxiv-cs-cv
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

The Token Games: Evaluating Language Model Reasoning with Puzzle Duels

DGX agent

arXiv:2602.17831v2 Announce Type: replace Abstract: Evaluating the reasoning capabilities of Large Language Models is increasingly challenging as models improve. Human curation of hard questions is hi

researcharxiv-cs-ai
19 May 2026
Model Releases

Threats to Arabic Handwriting Recognition: Investigating Black-Box Adversarial Attacks on embedded ConvNet models

DGX agent

arXiv:2605.18058v1 Announce Type: new Abstract: Arabic handwriting recognition (AHR) has made significant progress with deep learning models. AHR research has largely focused on performance, with secu

model-releasesarxiv-cs-cv
19 May 2026
Tutorials

Venom: A PyTorch Generative Modeling Toolkit

DGX agent

arXiv:2605.17605v1 Announce Type: new Abstract: Modern generative modeling has grown into a broad collection of related but often separately implemented paradigms, including denoising diffusion models

tutorialsarxiv-cs-lg
19 May 2026
Model Releases

WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform

DGX agent

arXiv:2605.17912v1 Announce Type: cross Abstract: World models have emerged as a central paradigm for embodied intelligence, enabling agents to predict action-conditioned future and reason about envir

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

A Cross-Modal Prompt Injection Attack against Large Vision-Language Models with Image-Only Perturbation

DGX agent

arXiv:2605.16090v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have emerged as a powerful paradigm for multimodal intelligence, but their growing deployment also expands the at

model-releasesarxiv-cs-cv
18 May 2026
Research

A numerical study into neural network surrogate model performance for uncertainty propagation

DGX agent

arXiv:2605.16078v1 Announce Type: cross Abstract: Neural network surrogate models have emerged as a promising approach to model solution fields for a wide variety of boundary value problems encountere

researcharxiv-cs-lg
18 May 2026
Tutorials

A Unified View of Score-Based and Drifting Models

DGX agent

arXiv:2603.07514v3 Announce Type: replace-cross Abstract: Drifting models train one-step generators by optimizing a kernel-induced mean-shift discrepancy between the data and model distributions, with

tutorialsarxiv-cs-ai
18 May 2026
Model Releases

Feedback World Model Enables Precise Guidance of Diffusion Policy

DGX agent

arXiv:2605.15705v1 Announce Type: cross Abstract: World models aim to improve robotic decision making by predicting the consequences of actions. However, in practice, their predictions often become un

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

DGX agent

arXiv:2509.12266v2 Announce Type: replace-cross Abstract: We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core c

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels

DGX agent

arXiv:2605.15208v1 Announce Type: cross Abstract: Large Language Models are routinely compressed via post-training quantization to reduce inference costs and memory footprint for cloud and edge deploy

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

A Large Language Model Based Pipeline for Review of Systems Entity Recognition from Clinical Notes

DGX agent

arXiv:2506.11067v3 Announce Type: replace Abstract: Objective: Develop a cost-effective, large language model (LLM)-based pipeline for automatically extracting Review of Systems (ROS) entities from cl

model-releasesarxiv-cs-cl
15 May 2026
Applications

An Interpretable Latency Model for Speculative Decoding in LLM Serving

DGX agent

arXiv:2605.15051v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model (LLM) inference by using a smaller draft model to propose multiple tokens that are verified b

applicationsarxiv-cs-lg
15 May 2026
Model Releases

Darwin Family: MRI-Trust-Weighted Evolutionary Merging for Training-Free Scaling of Language-Model Reasoning

DGX agent

arXiv:2605.14386v1 Announce Type: cross Abstract: We present Darwin Family, a framework for training-free evolutionary merging of large language models via gradient-free weight-space recombination. We

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Moral Susceptibility and Robustness under Persona Role-Play in Large Language Models

DGX agent

arXiv:2511.08565v3 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly operate in social contexts, motivating analysis of how they express and shift moral judgments. In th

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

NeuroMambaLLM: Dynamic Graph Learning of fMRI Functional Connectivity in Autistic Brains Using Mamba and Language Model Reasoning

DGX agent

arXiv:2602.13770v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong semantic reasoning across multimodal domains. However, their integration with graph-base

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

On the Cultural Anachronism and Temporal Reasoning in Vision Language Models

DGX agent

arXiv:2605.15071v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly applied to cultural heritage materials, from digital archives to educational platforms. This work ident

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Tokenizer Fertility and Zero-Shot Performance of Foundation Models on Ukrainian Legal Text: A Comparative Study

DGX agent

arXiv:2605.14890v1 Announce Type: new Abstract: Foundation models tokenize Ukrainian legal text with vastly different efficiency, yet no systematic comparison exists for this domain. We benchmark seve

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

AttenA+: Rectifying Action Inequality in Robotic Foundation Models

DGX agent

arXiv:2605.13548v1 Announce Type: cross Abstract: Existing robotic foundation models, while powerful, are predicated on an implicit assumption of temporal homogeneity: treating all actions as equally

model-releasesarxiv-cs-ai
14 May 2026
Research

Behavioral Geometric Supervision Aligns Video Foundation Models with Human Social Perception

DGX agent

arXiv:2510.01502v2 Announce Type: replace-cross Abstract: Current video foundation models, including the strongest self-supervised models such as V-JEPA2, fail to capture how humans organize social in

researcharxiv-cs-cv
14 May 2026
Model Releases

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

DGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

No One Knows the State of the Art in Geospatial Foundation Models

DGX agent

arXiv:2605.12678v1 Announce Type: new Abstract: Geospatial foundation models (GFMs) have been proposed as generalizable backbones for disaster response, land-cover mapping, food-security monitoring, a

model-releasesarxiv-cs-cv
14 May 2026
Research

Sample-Efficient Optimisation over the Outputs of Generative Models

DGX agent

arXiv:2509.23800v3 Announce Type: replace-cross Abstract: Modern generative AI models, such as diffusion and flow matching models, can sample from rich data distributions. However, many applications,

researcharxiv-cs-lg
14 May 2026
Applications

Bayesian Surrogate Training on Multiple Data Sources: A Hybrid Modeling Strategy

DGX agent

arXiv:2412.11875v3 Announce Type: replace-cross Abstract: Surrogate models are often used as computationally efficient approximations to complex simulation models, enabling tasks such as solving inver

applicationsarxiv-cs-lg
13 May 2026
Model Releases

Large-Small Model Collaboration for Farmland Semantic Change Detection

DGX agent

arXiv:2605.12282v1 Announce Type: new Abstract: Farmland Semantic Change Detection (SCD) is essential for cultivated land protection, yet existing benchmarks and models remain insufficient for fine-gr

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution

DGX agent

arXiv:2504.14129v4 Announce Type: replace Abstract: The challenge of tracing the source attribution of forged faces has gained significant attention due to the rapid advancement of generative models.

model-releasesarxiv-cs-cv
13 May 2026
Tutorials

Strategically Deceptive Model Deployment in Performative Prediction

DGX agent

arXiv:2506.09044v2 Announce Type: replace Abstract: Machine Learning systems are increasingly deployed in decision-making settings that shape user behavior and, in turn, the data on which future decis

tutorialsarxiv-cs-lg
13 May 2026
Research

Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives

DGX agent

arXiv:2406.05615v4 Announce Type: replace Abstract: Humans use multiple senses to comprehend the environment. Vision and language are two of the most vital senses since they allow us to easily communi

researcharxiv-cs-cl
13 May 2026
Research

Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing

DGX agent

arXiv:2605.10794v1 Announce Type: cross Abstract: Language models are deployed in settings that require compartmentalization: system prompts should not be disclosed, chain-of-thought reasoning is hidd

researcharxiv-cs-ai
12 May 2026
Model Releases

Language Models Without a Trainable Input Embedding Table: Learning from Fixed Minimal Binary Token Codes

DGX agent

arXiv:2605.09751v1 Announce Type: new Abstract: Trainable input embedding tables are a standard component of modern language models. We ask whether they are actually necessary at the input interface.

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Layer Collapse in Diffusion Language Models

DGX agent

arXiv:2605.06366v2 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as competitive alternatives to autoregressive (AR) language models, yet differences in their

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Learning Multi-Indicator Weights for Data Selection: A Joint Task-Model Adaptation Framework with Efficient Proxies

DGX agent

arXiv:2605.09665v1 Announce Type: cross Abstract: Data selection is a key component of efficient instruction tuning for large language models, as recent work has shown that data quality often matters

model-releasesarxiv-cs-ai
12 May 2026
Safety

Machine Unlearning on Pre-trained Models by Residual Feature Alignment Using LoRA

DGX agent

arXiv:2411.08443v2 Announce Type: replace-cross Abstract: Machine unlearning is an emerging technology that removes a subset of the training data from a trained model without significantly affecting t

safetyarxiv-cs-cv
12 May 2026
Model Releases

Model Capacity Determines Grokking through Competing Memorisation and Generalisation Speeds

DGX agent

arXiv:2605.09724v1 Announce Type: new Abstract: Existing accounts of grokking explain the phenomena in terms of mechanistic frameworks such as circuit efficiency or lazy-to-rich transitions. However,

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Preserve and Personalize: Personalized Text-to-Image Diffusion Models without Distributional Drift

DGX agent

arXiv:2505.19519v3 Announce Type: replace Abstract: Personalizing text-to-image diffusion models involves integrating novel visual concepts from a small set of reference images while retaining the mod

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari

DGX agent

arXiv:2605.08578v1 Announce Type: cross Abstract: Developing generalist systems that retain human-like data efficiency is a central challenge. While world models (WMs) offer a promising path, existing

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Recursive Language Models

DGX agent

arXiv:2512.24601v3 Announce Type: replace Abstract: We study allowing large language models (LLMs) to process arbitrarily long prompts through the lens of inference-time scaling. We propose Recursive

model-releasesarxiv-cs-ai
12 May 2026
Safety

TripleWin: Fixed-Point Equilibrium Pricing for Data-Model Coupled Markets

DGX agent

arXiv:2511.03368v2 Announce Type: replace Abstract: The rise of the machine learning (ML) model economy has intertwined markets for training datasets and pre-trained models. However, most pricing appr

safetyarxiv-cs-lg
12 May 2026
Model Releases

Adapting Vision-Language Models for Neutrino Event Classification in High-Energy Physics

DGX agent

arXiv:2509.08461v3 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated their remarkable capacity to process and reason over structured and unstruct

model-releasesarxiv-cs-ai
11 May 2026
Agents

AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites

DGX agent

arXiv:2605.06841v1 Announce Type: new Abstract: In model-based learning, the agent learns behaviors by simulating trajectories based on world model predictions. Standard world models typically learn a

agentsarxiv-cs-ai
11 May 2026
Model Releases

Detecting Distillation Data from Reasoning Models

DGX agent

arXiv:2510.04850v3 Announce Type: replace-cross Abstract: Reasoning distillation has emerged as a prevailing paradigm for transferring reasoning capabilities from large reasoning models to small langu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Evaluating Large Language Models in Scientific Discovery

DGX agent

arXiv:2512.15567v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet prevailing science benchmarks probe decontextualized knowledge and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

How Big Should a Wireless Foundation Model Be?

DGX agent

arXiv:2605.07266v1 Announce Type: cross Abstract: Wireless foundation models are rapidly emerging as a key enabler of AI-native communication systems, yet a fundamental question remains unanswered: ho

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models

DGX agent

arXiv:2605.06672v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning and reasoning-tuned models such as DeepSeek-R1 are commonly assumed to reduce shallow heuristic biases by thinking care

model-releasesarxiv-cs-ai
11 May 2026
Safety

Post-training makes large language models less human-like

DGX agent

arXiv:2605.07632v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavi

safetyarxiv-cs-ai
11 May 2026
Model Releases

Query-efficient model evaluation using cached responses

DGX agent

arXiv:2605.07096v1 Announce Type: cross Abstract: Evaluating a new model on an existing benchmark is often necessary to understand its behavior before deployment. For modern evaluation frameworks, gen

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

THINKSAFE: Self-Generated Safety Alignment for Reasoning Models

DGX agent

arXiv:2601.23143v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve remarkable performance by leveraging reinforcement learning (RL) on reasoning tasks to generate long chain-of-

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Emergent Hierarchical Structure in Large Language Models: An Information-Theoretic Framework for Multi-Scale Representation

DGX agent

arXiv:2505.18244v3 Announce Type: replace Abstract: Why do language models from different architecture families respond so differently to the same perturbation? We argue that the answer is not scale,

model-releasesarxiv-cs-cl
7 May 2026
← Previous
1…2627282930…1021
Next →