AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Safety

Do Models Fake Alignment Without Clear Consequences?

DGX agent

arXiv:2607.24758v1 Announce Type: new Abstract: Large language models are capable of recognizing evaluation contexts and altering their behavior to reflect evaluator expectations rather than typical d

safetyarxiv-cs-ai
29 Jul 2026
Safety

Faces of Fairness: Examining Bias in Facial Expression Recognition Datasets and Models

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2502.11049v3 Announce Type: replace Abstract: Automated Facial Expression Recognition (FER), involves two critical aspects: data and model design. Both significantly influence bias and fairness

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

Instruction-Tuned Models Locally Reuse Human Syntax More Than Humans Do

DGX agent

arXiv:2607.26015v1 Announce Type: new Abstract: Syntactic convergence (the tendency of speakers to adapt in language towards the grammatical profiles of their interlocutors) is a well-documented featu

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

SpecPrefetch: Parameter-Efficient Expert Prefetching for Sparse MoE Foundation Models

DGX agent

arXiv:2607.24787v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models expand foundation model capacity through conditional expert activation, but their full expert pools remain diffic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

A Scale-adaptive Vision Model Links C. elegans Neuronal Morphology to Behavior for Neurotoxicity Assessment

DGX agent

arXiv:2607.23183v1 Announce Type: cross Abstract: Neurological disorders are a leading cause of global disability and are increasingly linked to environmental chemical exposures. Yet neurotoxicity ass

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Chart Deception in Vision-Language Models: From Vulnerability to Mitigation

DGX agent

arXiv:2607.22600v1 Announce Type: new Abstract: Information visualizations are widely used to communicate patterns, trends, and outliers, yet deceptive design choices-such as truncated or inverted axe

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

IKS-Instruct: A 24,000-Example Multilingual Dataset for Teaching Language Models Indian Knowledge Systems

DGX agent

arXiv:2607.23322v1 Announce Type: new Abstract: Instruction tuning has become the standard method for adapting large language models to follow human intent, yet existing instruction datasets are domin

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization

DGX agent

arXiv:2607.22583v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved widespread adoption because of their strong reasoning and query-response capabilities. However, deploying the

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Rethinking the Generation Order of Block Diffusion Language Models

DGX agent

arXiv:2607.24306v1 Announce Type: new Abstract: Diffusion language models enable flexible arbitrary-order generation, but existing sampling methods are mostly designed for early masked diffusion model

researcharxiv-cs-cl
28 Jul 2026
Model Releases

StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design

DGX agent

arXiv:2607.22708v1 Announce Type: new Abstract: Deploying a vision-language model with full UI understanding on end devices has long been trapped between accuracy and efficiency: on one side is the ac

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research

DGX agent

arXiv:2607.24032v1 Announce Type: new Abstract: Generative-AI evaluations can become historical before publication, yet calendar age does not affect every conclusion equally. This paper has two linked

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Understanding Tone-Dependent Inference Cost in Large Language Models

DGX agent

arXiv:2607.23915v1 Announce Type: cross Abstract: We examine how prompt tone affects both accuracy of the LLM answers and inference cost as reflected in output-token consumption. Experiments were perf

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Atlas 2 -- Foundation models for clinical deployment

DGX agent

arXiv:2601.05148v2 Announce Type: replace Abstract: Pathology foundation models substantially advanced the possibilities in computational pathology --- yet tradeoffs in terms of performance, robustnes

researcharxiv-cs-cv
27 Jul 2026
Model Releases

DatedGPT: Preventing Lookahead Bias in Large Language Models with Time-Aware Pretraining

DGX agent

arXiv:2603.11838v2 Announce Type: replace Abstract: Large language models pretrained on internet-scale data risk lookahead bias in forecasting tasks, as they may have already seen the true outcome dur

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

GigaPath-Flash and GigaTIME-Flash: Efficient Pathology Foundation Models for Whole-Slide and Tumor Microenvironment Analysis

DGX agent

arXiv:2607.18218v2 Announce Type: replace-cross Abstract: Foundation models have emerged as a driving force in computational pathology, with the potential to transform cancer diagnosis, prognosis, and

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Training Large Language Models for Self-Explanation Faithfulness

DGX agent

arXiv:2607.21090v1 Announce Type: cross Abstract: We propose a Reinforcement Learning (RL) method to directly optimize the faithfulness of self-explanations - the extent to which a model's generated r

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Evaluating Large Language Models on Misconceptions in Multi-Turn Medical Conversations

DGX agent

arXiv:2607.12884v1 Announce Type: new Abstract: Patients seeking medical information often ask questions that embed incorrect assumptions or misconceptions. In such cases, safe medical communication r

model-releasesarxiv-cs-cl
15 Jul 2026
Safety

FlowWAM: Optical Flow as a Unified Action Representation for World Action Models

DGX agent

arXiv:2607.13017v1 Announce Type: cross Abstract: World Action Models (WAMs) are able to leverage pretrained video generators for both world modeling and action prediction. However, directly leveragin

safetyarxiv-cs-cv
15 Jul 2026
Model Releases

Scaling Point-in-Time Language Models

DGX agent

arXiv:2607.11889v1 Announce Type: cross Abstract: Large language models trained on unrestricted internet corpora inevitably embed information from the future, introducing lookahead bias that compromis

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

AtomBench: A Benchmarking Framework for Generative Crystal Reconstruction Models in Conventional Superconductors

DGX agent

arXiv:2510.16165v2 Announce Type: replace Abstract: A key question in benchmarking generative crystal reconstruction models is how the amount and type of crystallographic information provided to a gen

model-releasesarxiv-cs-lg
8 Jul 2026
Agents

MoWorld: A Flash World Model

DGX agent

arXiv:2607.06216v1 Announce Type: new Abstract: The future of World Models depends not only on scaling model capability, but also on scaling practicality and inference efficiency. High-frame-rate infe

agentsarxiv-cs-cv
8 Jul 2026
Model Releases

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models

DGX agent

arXiv:2506.07468v4 Announce Type: replace-cross Abstract: Conventional large language model (LLM) safety alignment relies on a reactive, disjoint loop: attackers exploit a static model, then defenders

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Continual Model Merging with Test-Time Adaptation for Whole-Slide Image Analysis

DGX agent

arXiv:2607.04755v1 Announce Type: new Abstract: Model merging offers a practical alternative to conventional continual learning by integrating independently fine-tuned models without retaining previou

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Is Your Benchmark Still Useful? Dynamic Benchmarking for Code Language Models

DGX agent

arXiv:2503.06643v2 Announce Type: replace-cross Abstract: In this paper, we tackle a critical challenge in model evaluation: how to keep code benchmarks useful when models might have already seen them

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Mask2Real-WM: Segmentation Masks as a Sim-to-Real Bridge for Controllable Dexterous World Models

DGX agent

arXiv:2607.04546v1 Announce Type: cross Abstract: Action-conditioned world models allow robots to predict the future consequences of candidate actions without additional physical interaction, supporti

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

The Remarkable Effectiveness of Providing AI Agents with Natural Language Tools: A Replication Study Validating NLT Performance Across 14 Models

DGX agent

arXiv:2607.03953v1 Announce Type: cross Abstract: This study independently replicates and extends the Natural Language Tools (NLT) framework of Johnson et al.~(2025), which questions the use of struct

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Discrete Diffusion Language Models for Interactive Radiology Report Drafting

DGX agent

arXiv:2607.01436v1 Announce Type: new Abstract: Diffusion language models, which generate text by denoising a token canvas bidirectionally instead of emitting tokens left to right, have become competi

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Gravity-Awareness: Deep Learning Models and LLM Simulation of Human Awareness in Altered Gravity

DGX agent

arXiv:2511.05536v2 Announce Type: replace-cross Abstract: Earth s gravity fundamentally shapes human behaviour. The brain encodes this force as an internal model of gravity, enabling the prediction an

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Liquid Latent State Dynamics for Interpretable Turbofan Degradation Modeling

DGX agent

arXiv:2607.01986v1 Announce Type: new Abstract: Multivariate time-series models for prognostics are often evaluated by point prediction accuracy, yet their internal states rarely expose a coherent deg

model-releasesarxiv-cs-lg
3 Jul 2026
Tutorials

Harnessing the Latent Space: From Steering Vectors to Model Calibrators for Control and Trust

DGX agent

arXiv:2607.00083v1 Announce Type: cross Abstract: Language models have changed from unreliable text generators to highly-capable large models with trillions of parameters. Capability increases come ha

tutorialsarxiv-cs-ai
2 Jul 2026
Research

How Post-Training Shapes Biological Reasoning Models

DGX agent

arXiv:2606.16517v2 Announce Type: replace Abstract: Scientific reasoning models for biology combine language models with foundation models trained on multimodal biological data, including DNA, RNA, an

researcharxiv-cs-lg
1 Jul 2026
Research

Concentration bounds on response-based vector embeddings of black-box generative models

DGX agent

arXiv:2511.08307v2 Announce Type: replace-cross Abstract: Generative models, such as large language models or text-to-image diffusion models, can generate relevant responses to user-given queries. Res

researcharxiv-cs-lg
30 Jun 2026
Model Releases

DNA Language Models: An Assessment of Pre-Training for Fine-Tuning Tasks

DGX agent

arXiv:2606.30140v1 Announce Type: cross Abstract: Recent breakthroughs in foundation models and Large Language Models (LLMs) have introduced new opportunities for studying and decoding genomic sequenc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

FlipGuard: Defending Large Language Models Against Quantization-Conditioned Backdoor Attacks

DGX agent

arXiv:2606.28962v1 Announce Type: cross Abstract: Model quantization is essential for the efficient deployment of Large Language Models (LLMs), but introduces a critical vulnerability: Quantization-Co

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Flow Matching in Feature Space for Stochastic World Modeling

DGX agent

arXiv:2606.29059v1 Announce Type: cross Abstract: World modeling requires forecasting uncertain futures while preserving information useful for downstream perception. Existing visual world models ofte

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Fuzzing Large Language Models to Elicit Hidden Behaviours

DGX agent

arXiv:2606.29646v1 Announce Type: cross Abstract: Sleeper agents are the canonical model organism of deception: models trained to behave normally but to emit an unsafe behaviour on a specific trigger.

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Little Brains, Big Feats: Exploring Compact Language Models

DGX agent

arXiv:2606.30062v1 Announce Type: cross Abstract: While large language models have been dominating the research landscape recently, small language models remain highly relevant across various domains;

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SurgVLA-Bench: Towards Evaluating Vision-Language-Action Models for Laparoscopic Surgical Robotics

DGX agent

arXiv:2606.29247v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models represent a promising direction for embodied intelligence in surgical robotics. Despite the prevalence of VLA benchm

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Revisiting Performance Claims for Chest X-Ray Models Using Clinical Context

DGX agent

arXiv:2509.19671v3 Announce Type: replace Abstract: Public datasets of Chest X-Rays (CXRs) have long been a popular benchmark for developing machine learning (ML) computer vision models in healthcare.

model-releasesarxiv-cs-lg
29 Jun 2026
Agents

EvoOptiGraph: Weakness-Driven Coevolution via Graph-Based Structural Generation for Optimization Modeling

DGX agent

arXiv:2606.26578v1 Announce Type: new Abstract: Automating optimization modeling from natural language with large language models (LLMs) faces two key challenges. First, training corpora lack structur

agentsarxiv-cs-ai
26 Jun 2026
Local Ai

Not All Actions Are Equal: Rethinking Conditioning for Dexterous World Model

DGX agent

arXiv:2606.27325v1 Announce Type: new Abstract: Recent advances in action-conditioned world models show promising progress in modeling complex interactions and forecasting future states under diverse

local-aiarxiv-cs-cv
26 Jun 2026
Tutorials

Did Models Learn Sufficiently? Attribution-Guided Training via Subset-Selected Counterfactual Augmentation

DGX agent

arXiv:2511.12100v2 Announce Type: replace Abstract: In current visual model training, models often rely on only limited sufficient causes for their predictions, which makes them sensitive to distribut

tutorialsarxiv-cs-cv
25 Jun 2026
Model Releases

CAVEWOMAN: How Large Language Models Behave Under Linguistic Input and Output Compression

DGX agent

arXiv:2606.24083v1 Announce Type: cross Abstract: 'Talk short. Drop grammar. Save token.' This caveman style is widely promoted as a way to cut inference cost, but whether it actually saves anything d

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Trimming the Long-Tail of Visual World Modeling Evaluation

DGX agent

arXiv:2606.24256v1 Announce Type: new Abstract: Physical interactions follow a long-tailed distribution: a set of common and regular interactions dominates human experience and visual data, while a br

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model

DGX agent

arXiv:2606.22317v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is widely viewed as a promising path toward continuously improving large language models. Recent w

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

GraphPFN: A Prior-Data Fitted Graph Foundation Model

DGX agent

arXiv:2509.21489v3 Announce Type: replace Abstract: Graph foundation models face several fundamental challenges including transferability across diverse domains and data scarcity, which calls into que

applicationsarxiv-cs-lg
23 Jun 2026
Research

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding

DGX agent

arXiv:2606.20726v1 Announce Type: new Abstract: We introduce a compact empirical model that quantifies how answer accuracy degrades as a function of frame budget B and temporal distance D in long vide

researcharxiv-cs-cv
23 Jun 2026
Research

PACT: Preserving Anchored Cores in Task-vectors for Model Merging

DGX agent

arXiv:2606.18627v2 Announce Type: replace Abstract: Model merging has emerged as a training-free alternative to multi-task learning, aiming to combine multiple task-specific fine-tuned models into a s

researcharxiv-cs-lg
23 Jun 2026
← Previous
1…1617181920…1012
Next →