AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Research

How Post-Training Shapes Biological Reasoning Models

DGX agent

arXiv:2606.16517v2 Announce Type: replace Abstract: Scientific reasoning models for biology combine language models with foundation models trained on multimodal biological data, including DNA, RNA, an

researcharxiv-cs-lg
1 Jul 2026
Research

Concentration bounds on response-based vector embeddings of black-box generative models

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2511.08307v2 Announce Type: replace-cross Abstract: Generative models, such as large language models or text-to-image diffusion models, can generate relevant responses to user-given queries. Res

researcharxiv-cs-lg
30 Jun 2026
Model Releases

DNA Language Models: An Assessment of Pre-Training for Fine-Tuning Tasks

DGX agent

arXiv:2606.30140v1 Announce Type: cross Abstract: Recent breakthroughs in foundation models and Large Language Models (LLMs) have introduced new opportunities for studying and decoding genomic sequenc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

FlipGuard: Defending Large Language Models Against Quantization-Conditioned Backdoor Attacks

DGX agent

arXiv:2606.28962v1 Announce Type: cross Abstract: Model quantization is essential for the efficient deployment of Large Language Models (LLMs), but introduces a critical vulnerability: Quantization-Co

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Flow Matching in Feature Space for Stochastic World Modeling

DGX agent

arXiv:2606.29059v1 Announce Type: cross Abstract: World modeling requires forecasting uncertain futures while preserving information useful for downstream perception. Existing visual world models ofte

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Fuzzing Large Language Models to Elicit Hidden Behaviours

DGX agent

arXiv:2606.29646v1 Announce Type: cross Abstract: Sleeper agents are the canonical model organism of deception: models trained to behave normally but to emit an unsafe behaviour on a specific trigger.

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Little Brains, Big Feats: Exploring Compact Language Models

DGX agent

arXiv:2606.30062v1 Announce Type: cross Abstract: While large language models have been dominating the research landscape recently, small language models remain highly relevant across various domains;

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SurgVLA-Bench: Towards Evaluating Vision-Language-Action Models for Laparoscopic Surgical Robotics

DGX agent

arXiv:2606.29247v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models represent a promising direction for embodied intelligence in surgical robotics. Despite the prevalence of VLA benchm

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Revisiting Performance Claims for Chest X-Ray Models Using Clinical Context

DGX agent

arXiv:2509.19671v3 Announce Type: replace Abstract: Public datasets of Chest X-Rays (CXRs) have long been a popular benchmark for developing machine learning (ML) computer vision models in healthcare.

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

I put together a new article on setting up local coding agents with open-weight models. Everything runs 100% locally. I thought it might be …

DGX agent

I put together a new article on setting up local coding agents with open-weight models. Everything runs 100% locally. I thought it might be useful putting this together because many people asked me ab

model-releasessebastian-raschka--x
27 Jun 2026
Agents

EvoOptiGraph: Weakness-Driven Coevolution via Graph-Based Structural Generation for Optimization Modeling

DGX agent

arXiv:2606.26578v1 Announce Type: new Abstract: Automating optimization modeling from natural language with large language models (LLMs) faces two key challenges. First, training corpora lack structur

agentsarxiv-cs-ai
26 Jun 2026
Local Ai

Not All Actions Are Equal: Rethinking Conditioning for Dexterous World Model

DGX agent

arXiv:2606.27325v1 Announce Type: new Abstract: Recent advances in action-conditioned world models show promising progress in modeling complex interactions and forecasting future states under diverse

local-aiarxiv-cs-cv
26 Jun 2026
Tutorials

Did Models Learn Sufficiently? Attribution-Guided Training via Subset-Selected Counterfactual Augmentation

DGX agent

arXiv:2511.12100v2 Announce Type: replace Abstract: In current visual model training, models often rely on only limited sufficient causes for their predictions, which makes them sensitive to distribut

tutorialsarxiv-cs-cv
25 Jun 2026
Model Releases

CAVEWOMAN: How Large Language Models Behave Under Linguistic Input and Output Compression

DGX agent

arXiv:2606.24083v1 Announce Type: cross Abstract: 'Talk short. Drop grammar. Save token.' This caveman style is widely promoted as a way to cut inference cost, but whether it actually saves anything d

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Trimming the Long-Tail of Visual World Modeling Evaluation

DGX agent

arXiv:2606.24256v1 Announce Type: new Abstract: Physical interactions follow a long-tailed distribution: a set of common and regular interactions dominates human experience and visual data, while a br

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model

DGX agent

arXiv:2606.22317v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is widely viewed as a promising path toward continuously improving large language models. Recent w

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

GraphPFN: A Prior-Data Fitted Graph Foundation Model

DGX agent

arXiv:2509.21489v3 Announce Type: replace Abstract: Graph foundation models face several fundamental challenges including transferability across diverse domains and data scarcity, which calls into que

applicationsarxiv-cs-lg
23 Jun 2026
Research

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding

DGX agent

arXiv:2606.20726v1 Announce Type: new Abstract: We introduce a compact empirical model that quantifies how answer accuracy degrades as a function of frame budget B and temporal distance D in long vide

researcharxiv-cs-cv
23 Jun 2026
Research

PACT: Preserving Anchored Cores in Task-vectors for Model Merging

DGX agent

arXiv:2606.18627v2 Announce Type: replace Abstract: Model merging has emerged as a training-free alternative to multi-task learning, aiming to combine multiple task-specific fine-tuned models into a s

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Sub-Billion, Super-Frontier: Small Language Models Rival Zero-Shot Frontier LLMs on General and Literary Relation Extraction

DGX agent

arXiv:2606.22606v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong relation extraction (RE), but their computational demands and reliance on proprietary APIs limit deploymen

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer

DGX agent

arXiv:2511.22699v4 Announce Type: replace Abstract: The landscape of high-performance image generation models is currently dominated by proprietary systems, such as Nano Banana Pro and Seedream 4.0. L

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Ai2 just released TMax 27B on Hugging Face A 27B terminal agent that hits 42.7% on Terminal Bench 2.0, rivaling models 40× its size.

DGX agent

AI2 released TMax 27B, a 27 billion parameter terminal agent model available on Hugging Face that achieves 42.7% performance on Terminal Bench 2.0, matching the capabilities of much larger models desp

model-releasesclem-delangue--x
22 Jun 2026
Model Releases

Overcoming State Inertia in Full-Duplex Spoken Language Models via Activation Steering

DGX agent

arXiv:2606.11386v1 Announce Type: cross Abstract: Full-duplex spoken language models (FD-SLMs) enable seamless speech interaction by allowing models to listen and speak simultaneously, yet the interna

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation

DGX agent

arXiv:2606.11270v1 Announce Type: cross Abstract: Distillation of a language model intended to transfer benign behavior to a student model may also transfer undesirable characteristics, if they are pr

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Domain Adapted Large Language Models for Additive Manufacturing

DGX agent

arXiv:2603.22017v2 Announce Type: replace Abstract: This work presents a collection of multi-modal domain adapted large language models built upon the instruction tuned variants of open weight models

model-releasesarxiv-cs-lg
10 Jun 2026
Research

One Lens, Many Worlds : A Capability-Typed Interface for World-Model Interpretability

DGX agent

arXiv:2606.09936v1 Announce Type: cross Abstract: World models are now built on substantially different computational substrates. Latent recurrent state-space models such as PlaNet and the Dreamer fam

researcharxiv-cs-ai
10 Jun 2026
Model Releases

PhantomBench: Benchmarking the Non-existential Threat of Language Models

DGX agent

arXiv:2606.11105v1 Announce Type: cross Abstract: Hallucinations, where language models (LMs) generate factually ungrounded responses, pose serious risks, as users tend to blindly rely on them. This i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

When RL Fails after SFT: Rejuvenating Model Plasticity for Robust SFT-to-RL Handoff

DGX agent

arXiv:2606.09932v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) has become a standard pipeline for Large Language Model (LLM) post-training. SFT

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

DriveReward: A Comprehensive Dataset and Generative Vision-Language Reward Model for Autonomous Driving

DGX agent

arXiv:2606.08525v1 Announce Type: new Abstract: Reward models play a pivotal role in reinforcement learning (RL) and multi-modal trajectory selection for autonomous driving. However, acquiring such re

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

NutriMLLM: Multimodal Large Language Models for Dietary Micronutrient Analysis

DGX agent

arXiv:2606.08948v1 Announce Type: cross Abstract: Comprehensive estimation of dietary micronutrients from food images could improve clinical nutrition care, but training such models requires large mul

model-releasesarxiv-cs-ai
9 Jun 2026
Applications

Design Once, Deploy at Scale: Template-Driven ML Development for Large Model Ecosystems

DGX agent

arXiv:2603.24963v3 Announce Type: replace Abstract: Modern computational advertising platforms typically rely on recommendation systems to predict user responses, such as click-through rates, conversi

applicationsarxiv-cs-ai
8 Jun 2026
Model Releases

Machine Learning for Electron-Scale Turbulence Modeling in W7-X

DGX agent

arXiv:2511.04567v2 Announce Type: replace-cross Abstract: Constructing reduced models for turbulent transport is essential for accelerating profile predictions and enabling many-query tasks such as pa

model-releasesarxiv-cs-lg
8 Jun 2026
Research

Understanding Generative Recommendation with Semantic IDs from a Model-scaling View

DGX agent

arXiv:2509.25522v3 Announce Type: replace Abstract: Recent advancements in generative models have allowed the emergence of a promising paradigm for recommender systems (RS), known as Generative Recomm

researcharxiv-cs-ai
8 Jun 2026
Model Releases

VLA-JEPA just dropped in LeRobot 🤖 What makes this model special is that it does not just learn what action to take from a given observatio…

DGX agent

VLA-JEPA just dropped in LeRobot 🤖 What makes this model special is that it does not just learn what action to take from a given observation, it also leverages a JEPA world model to learn action-relev

model-releasesclem-delangue--x
6 Jun 2026
Model Releases

Dream.exe: Can Video Generation Models Dream Executable Robot Manipulation?

DGX agent

arXiv:2606.04811v1 Announce Type: new Abstract: Video generation models have made impressive strides in synthesizing visually compelling content, yet their outputs remain confined to the virtual domai

model-releasesarxiv-cs-cv
4 Jun 2026
Safety

Generalization of World Models under Environmental Variability for Vision-based Quadrotor Navigation

DGX agent

arXiv:2606.05015v1 Announce Type: new Abstract: World models, learned generative models that predict how an environment evolves, have become a promising tool for sample-efficient robot learning. Yet h

safetyarxiv-cs-ro
4 Jun 2026
Model Releases

Geometry-Aware Hallucination Detection in Large Language Models

DGX agent

arXiv:2601.06196v3 Announce Type: replace-cross Abstract: Large language models (LLMs) frequently generate factually incorrect or unsupported content, commonly referred to as hallucinations. Prior wor

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model

DGX agent

arXiv:2512.21917v3 Announce Type: replace-cross Abstract: Policy alignment to preference data typically assumes a known link function between observed preferences and latent rewards (e.g., Bradley-Ter

safetyarxiv-cs-ai
4 Jun 2026
Research

Token Rankings are Unforgeable Language Model Signatures

DGX agent

arXiv:2606.04459v1 Announce Type: cross Abstract: Language model parameters are known to impose unique (to each model) geometric constraints on their logit outputs, which serves as a signature that id

researcharxiv-cs-ai
4 Jun 2026
Research

UniCanvas: A Diffusion-base Unified Model for Text-in-Image Joint Generation

DGX agent

arXiv:2606.04264v1 Announce Type: new Abstract: Recent years have seen remarkable progress in unified vision-language models handling both multimodal understanding and generation within a single archi

researcharxiv-cs-cv
4 Jun 2026
Model Releases

ChatHealthAI: Aligning Electronic Health Record Representations with Large Language Models for Grounded Clinical Reasoning

DGX agent

arXiv:2606.02802v1 Announce Type: new Abstract: Large language models (LLMs) exhibit strong natural-language reasoning abilities for clinical decision support, but struggle to effectively model struct

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Hybrid Dynamics Modeling for a Flexible 2-DoF Robotic Arm

DGX agent

arXiv:2606.02969v1 Announce Type: new Abstract: This paper examines three approaches for modeling the dynamics of a flexible-link 2-DoF robotic arm to address unmodeled dynamics not captured by rigid-

model-releasesarxiv-cs-ro
3 Jun 2026
Research

Large Byte Model: Teaching Language Models About Compiled Code

DGX agent

arXiv:2606.02834v1 Announce Type: cross Abstract: Malware analysis starts with the raw bytes of an executable program, and tools to 'lift' these to higher-level representations, such as assembly, are

researcharxiv-cs-ai
3 Jun 2026
Model Releases

3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code

DGX agent

arXiv:2606.01057v1 Announce Type: cross Abstract: Procedural 3D modeling through code is emerging as a versatile paradigm, offering deterministic, engine-ready, and precisely editable assets that neur

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Large Electron Model: A Universal Ground State Predictor

DGX agent

arXiv:2603.02346v2 Announce Type: replace-cross Abstract: We introduce Large Electron Model, a single neural network model that produces variational wavefunctions of interacting electrons over the ent

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LL-Bench: Rethinking Low-Level Vision Evaluation in the Era of Large-Scale Generative Models

DGX agent

arXiv:2606.02535v1 Announce Type: new Abstract: Large-scale generative models have demonstrated remarkable capabilities across image generation and editing tasks. However, their performance in low-lev

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

On the Generalization Gap in Self-Evolving Language Model Reasoning

DGX agent

arXiv:2606.01075v1 Announce Type: new Abstract: Recent work suggests that large language models (LLMs) can improve through self-evolution (SE), using supervision signals generated by the model itself.

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Saliency-Aware Model Merging

DGX agent

arXiv:2606.00511v1 Announce Type: cross Abstract: Model merging aims to consolidate multiple task-specific models fine-tuned on different datasets into a unified architecture that performs cross-domai

model-releasesarxiv-cs-cv
2 Jun 2026
← Previous
1…2728293031…1247
Next →