AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Safety

ReWorld: Learning Better Representations for World Action Models

DGX agent

arXiv:2606.27504v1 Announce Type: new Abstract: World Action Models (WAMs) model future environment evolution under action conditioning, offering a scalable paradigm for autonomous driving. However, e

safetyarxiv-cs-cv
29 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision

DGX agent

arXiv:2512.21970v2 Announce Type: replace Abstract: While Vision-Language-Action (VLA) models excel in generalist manipulation, they often lack fine-grained spatial awareness and show limited viewpoin

model-releasesarxiv-cs-ro
29 Jun 2026
Research

The Weakest Link Tells It All: Outcome-Supervised Process Reward Modeling via Learnable Credit Assignment

DGX agent

arXiv:2606.27739v1 Announce Type: new Abstract: Process reward models (PRMs) enhance the reasoning capabilities of large language models (LLMs) by providing fine-grained feedback, yet training PRMs ty

researcharxiv-cs-lg
29 Jun 2026
Model Releases

Your AI Travel Agent Would Book You a Bullfight: An Agentic Benchmark for Implicit Animal Welfare in Frontier AI Models

DGX agent

arXiv:2606.18142v3 Announce Type: replace Abstract: AI agents are moving from advisors to actors, booking travel, planning menus, and running procurement on behalf of users. Existing benchmarks for AI

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Know2Guess: A Contamination-Aware Multi-Zone Benchmark for Knowledge-Boundary Evaluation in Large Language Models

DGX agent

arXiv:2606.26101v1 Announce Type: cross Abstract: Reliable evaluation of large language models should separate supported answering from unsupported guessing without conflating either with data contami

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

DGX agent

arXiv:2606.26968v1 Announce Type: new Abstract: Speech-capable models are increasingly deployed in real-world applications across languages. Yet their safety and fairness beyond English settings and u

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

DGX agent

arXiv:2510.09976v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and pi_0 have shown strong generalization by leveraging large-scale demonstrations, yet t

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

DGX agent

arXiv:2606.18430v2 Announce Type: replace Abstract: Statistical watermarks help organizations attribute large language model (LLM) outputs, yet existing detectors often struggle when watermark signals

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Are Tabular Foundation Models Robust to Realistic Query Distribution Shifts in Microbiome Data?

DGX agent

arXiv:2606.24995v1 Announce Type: new Abstract: Tabular foundation models (TFMs) achieve strong performance on microbiome abundance data, yet their robustness under realistic distribution shift remain

model-releasesarxiv-cs-lg
25 Jun 2026
Research

Certification of Machine Learning Models via Directional Sharpness

DGX agent

arXiv:2606.25004v1 Announce Type: new Abstract: In machine learning, model certification has been identified as an important method for gaining assurance about a model's trustworthiness and quality. A

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Efficient Remote Sensing Instance Segmentation with Linear-Time State Space Distilled Visual Foundation Models

DGX agent

arXiv:2606.25324v1 Announce Type: new Abstract: The computational complexity of Transformers scales quadratically with the number of tokens, which significantly constrains the efficiency of vision mod

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

EPTS: Elastic Post-Training Sparsity for Efficient Large Language Model Compression

DGX agent

arXiv:2606.25285v1 Announce Type: new Abstract: Post-Training Sparsity (PTS) has emerged as a crucial paradigm for compressing Large Language Models to facilitate efficient deployment on resource-cons

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Perfect Detection, Failed Control: The Geometry of Knowing vs. Steering in Language Models

DGX agent

arXiv:2606.24952v1 Announce Type: new Abstract: A central aspiration of mechanistic interpretability is controllability: if we know where a behavior is represented in a model's activations, we should

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

Quantization Inflates Reasoning: Token Inflation as a Hidden Cost of Low-Bit Reasoning Models

DGX agent

arXiv:2606.25519v1 Announce Type: cross Abstract: Quantization is widely used to reduce the inference cost of large language models, but its effect on reasoning models is not fully captured by final-a

agentsarxiv-cs-lg
25 Jun 2026
Model Releases

Robustness assessment of large audio language models in multiple-choice evaluation

DGX agent

arXiv:2510.04584v2 Announce Type: replace Abstract: Recent advances in large audio language models (LALMs) have primarily been assessed using a multiple-choice question answering (MCQA) framework. How

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation

DGX agent

arXiv:2510.09036v2 Announce Type: replace Abstract: Learned world models hold significant potential as neural simulators for robotic manipulation. However, prevalent 2D video-based models inherently l

agentsarxiv-cs-ro
25 Jun 2026
Model Releases

Same Evidence, Different Answer: Auditing Order Sensitivity in Multimodal Large Language Models

DGX agent

arXiv:2606.26079v1 Announce Type: new Abstract: Standard benchmarks for multimodal large language models (MLLMs) score each item on one canonical ordering and miss whether order-irrelevant shuffling c

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Small Initialization Matters for Large Language Models

DGX agent

arXiv:2606.17945v2 Announce Type: replace Abstract: Large language models provide a tractable system for asking how intelligence itself emerges, rather than only how LLMs can be engineered. Although p

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models

DGX agent

arXiv:2606.24962v1 Announce Type: new Abstract: Recent progress in large-scale sequence modeling has shown that a single model can learn useful representations across highly diverse data distributions

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability

DGX agent

arXiv:2606.24589v1 Announce Type: new Abstract: Scaling adversarial evaluation of large language models requires both a method for generating hard inputs and a reliable way to confirm that resulting f

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Business as Rulesual: A Benchmark and Framework for Business Rule Flow Modeling with LLMs

DGX agent

arXiv:2505.18542v4 Announce Type: replace Abstract: Extracting structured procedural knowledge from unstructured business documents is a critical yet unresolved bottleneck in process automation. While

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Can Scale Save Us From Plasticity Loss in Large Language Models?

DGX agent

arXiv:2606.24752v1 Announce Type: new Abstract: The loss of plasticity - the ability of a network to learn new information after having already learned older information - is a fundamental challenge i

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Extended pseudo-spectral physics-informed neural networks for phase-field models

DGX agent

arXiv:2606.24660v1 Announce Type: cross Abstract: Phase-field models play a central role in the continuum description of phase separation, in which the bulk free-energy density and the interfacial thi

model-releasesarxiv-cs-lg
24 Jun 2026
Applications

HyMaTE: A Hybrid Mamba and Transformer Model for EHR Representation Learning

DGX agent

arXiv:2509.24118v2 Announce Type: replace Abstract: Electronic health Records (EHRs) have become a cornerstone in modern-day healthcare. They are a crucial part for analyzing the progression of patien

applicationsarxiv-cs-lg
24 Jun 2026
Model Releases

Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients

DGX agent

arXiv:2606.24774v1 Announce Type: new Abstract: Vision-Language Large Models (VLLMs) trained on massive crawled corpora raise pressing copyright and data-provenance concerns. These concerns are partic

model-releasesarxiv-cs-cv
24 Jun 2026
Research

Sentence-Level Contextual Entrainment in Large Language Models

DGX agent

arXiv:2606.24077v1 Announce Type: new Abstract: Contextual entrainment, which is a newly discovered phenomenon in large language models (LLMs), refers to the tendency of a model to assign higher proba

researcharxiv-cs-cl
24 Jun 2026
Model Releases

Transformer-Based Language Models Across Domain Verticals: Architectures, Applications and Critical Assessment

DGX agent

arXiv:2606.24331v1 Announce Type: new Abstract: Transformer-based language models have become the default substrate for natural language processing and the pace of new releases has made it hard for pr

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

A-Evolve-Training: Autonomous Post-Training of a 30B Model

DGX agent

arXiv:2606.20657v1 Announce Type: cross Abstract: Post-training a frontier model is normally weeks of human work: proposing data and recipe changes, launching runs, reading evals, deciding what to kee

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

A Linear Fractional Transformation Model and Calibration Method for Light Field Camera

DGX agent

arXiv:2511.03962v2 Announce Type: replace Abstract: Accurate intrinsic calibration is a crucial yet challenging prerequisite for 3D reconstruction using light field cameras. Existing calibration model

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Attacking the Trusted Imagination: Oracle-Level Integrity Attacks on Imagine-then-Act World Models

DGX agent

arXiv:2606.22966v1 Announce Type: new Abstract: Many recent vision-language-action (VLA) policies adopt an imagine-then-act design. A world-action model (WAM) first imagines a short future as a latent

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models

DGX agent

arXiv:2606.21645v1 Announce Type: cross Abstract: Large language models (LLMs) can readily reproduce conventional expressions, yet their ability to model gradient frequency distributions remains under

safetyarxiv-cs-lg
23 Jun 2026
Tutorials

Diffusion-Driven State Space Models

DGX agent

arXiv:2606.21036v1 Announce Type: cross Abstract: In many domains, practitioners seek models that produce accurate forecasts while faithfully capturing latent system dynamics. Existing approaches typi

tutorialsarxiv-cs-lg
23 Jun 2026
Model Releases

Evaluating Large Language Models for Hausa and Fongbe Machine Translation: Benchmarks, Failures, and Metric Reliability

DGX agent

arXiv:2606.22269v1 Announce Type: cross Abstract: We investigate the translation quality of current large language models (LLMs) for English-to-Hausa and English-to-Fongbe - two typologically distinct

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

How Well Do Self-Supervised Speech Models Encode Age and Gender in Children's Speech? A Layer-Wise Analysis Across Multiple Architectures

DGX agent

arXiv:2606.22177v1 Announce Type: cross Abstract: Self-supervised learning (SSL) models have become a central component of modern speech processing systems, as they enable the learning of rich acousti

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

L20-Edu-135M: An Auditable Single-GPU Study of Data-Efficient Small Language Modeling

DGX agent

arXiv:2606.22189v1 Announce Type: new Abstract: Small language models are cheap to serve and feasible on local hardware, but strong public 135M-class systems are commonly trained with hundreds of bill

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Learning-Based Modeling of Soft Robots via Cosserat Rod Theory

DGX agent

arXiv:2606.20958v1 Announce Type: new Abstract: Modeling soft robot dynamics is challenging due to their continuum structure and typically nonlinear dynamics. Creating models based on first-order prin

researcharxiv-cs-ro
23 Jun 2026
Model Releases

LIBERO-Safety: A Comprehensive Benchmark for Physical and Semantic Safety in Vision-Language-Action Models

DGX agent

arXiv:2606.23686v1 Announce Type: new Abstract: Despite the impressive manipulation capabilities of Vision-Language-Action (VLA) models, their operational safety under strict constraints remains large

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Load Testing for Machine Learning Model Serving Systems at Scale

DGX agent

arXiv:2606.22013v1 Announce Type: new Abstract: Machine learning (ML) model serving has become a dominant consumer of GPU infrastructure, yet capacity planning in these systems remains largely ad hoc.

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models

DGX agent

arXiv:2509.23729v3 Announce Type: replace Abstract: Large Language Models (LLMs) with multimodal capabilities have revolutionized vision-language tasks, but their deployment often requires huge memory

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Model Inversion meets Cryptographic Fuzzy Extractors

DGX agent

arXiv:2510.25687v3 Announce Type: replace-cross Abstract: Model inversion attacks pose an open challenge to privacy-sensitive applications that use machine learning (ML) models. For example, face auth

researcharxiv-cs-lg
23 Jun 2026
Safety

PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

DGX agent

arXiv:2606.18375v2 Announce Type: replace Abstract: World foundation models (WFMs) are powerful simulators, yet they predominantly operate in a single-view setting and lack the multi-view 3D consisten

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

Post-Training Speech Enhancement Language Models with Perceptual Rewards

DGX agent

arXiv:2606.21458v1 Announce Type: new Abstract: Speech enhancement language models achieve strong results when trained on discrete audio tokens, but their optimization relies on token-level cross-entr

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Residue-Level Attributions in Protein Language Models Do Not Recover Allergen Epitopes

DGX agent

arXiv:2606.22181v1 Announce Type: new Abstract: Deep allergenicity classifiers are increasingly used in safety screening of novel foods, and recent protein language models have substantially improved

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Translating Inference-Time Control to Radiology Vision-Language Models: Activation Steering for Pneumonia Classification on Chest X-rays

DGX agent

arXiv:2606.20852v1 Announce Type: new Abstract: Inference-time engineering can alter model behavior without fine-tuning. However, its utility for improving diagnostic performance in medical vision-lan

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training

DGX agent

arXiv:2508.03058v2 Announce Type: replace Abstract: Reinforcement Learning (RL) in real-world environments often suffers from ambiguous or incomplete reward supervision, which undermines policy stabil

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

OpenMedReason: Scientific Reasoning Supervision for Medical Vision-Language Models

DGX agent

arXiv:2606.12169v1 Announce Type: cross Abstract: High-stakes clinical use of large vision-language models (LVLMs) requires reasoning that is grounded in visual evidence and clinical knowledge, not ju

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task

DGX agent

arXiv:2606.11830v1 Announce Type: new Abstract: Background. Large language models and AI agents are increasingly used to support biomedical research, but native model outputs may omit key analytical s

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

TimeRouter: Efficient and Adaptive Routing of Time-Series Foundation Models

DGX agent

arXiv:2606.11625v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) are increasingly explored as predictive experts within emerging agentic time-series systems. However, TSFMs exhibi

model-releasesarxiv-cs-lg
11 Jun 2026
← Previous
1…6162636465…1021
Next →