AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,489 results
Model Releases

StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision

DGX agent

arXiv:2512.21970v2 Announce Type: replace Abstract: While Vision-Language-Action (VLA) models excel in generalist manipulation, they often lack fine-grained spatial awareness and show limited viewpoin

model-releasesarxiv-cs-ro
29 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

The Weakest Link Tells It All: Outcome-Supervised Process Reward Modeling via Learnable Credit Assignment

DGX agent

arXiv:2606.27739v1 Announce Type: new Abstract: Process reward models (PRMs) enhance the reasoning capabilities of large language models (LLMs) by providing fine-grained feedback, yet training PRMs ty

researcharxiv-cs-lg
29 Jun 2026
Model Releases

Your AI Travel Agent Would Book You a Bullfight: An Agentic Benchmark for Implicit Animal Welfare in Frontier AI Models

DGX agent

arXiv:2606.18142v3 Announce Type: replace Abstract: AI agents are moving from advisors to actors, booking travel, planning menus, and running procurement on behalf of users. Existing benchmarks for AI

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Grok 4.5, based on our 1.5T V9 foundation model, with Cursor data added in supplemental training, is now in private beta at SpaceX & Tesla. …

DGX agent

Grok 4.5, based on our 1.5T V9 foundation model, with Cursor data added in supplemental training, is now in private beta at SpaceX & Tesla. Early evals show performance close to, perhaps exceeding Opu

model-releaseselon-musk--x
28 Jun 2026
Model Releases

And the idea that frontier open weights models will continue to be released was fragile even before the current de facto licensing regime. h…

DGX agent

And the idea that frontier open weights models will continue to be released was fragile even before the current de facto licensing regime. https://x.com/emollick/status/2067669551685218638?s=20 Is the

model-releasesethan-mollick--x
27 Jun 2026
Model Releases

VCs are now sharing screenshots in group chats of Claude discouraging investment in open-source AI infra startups and models. Obviously ther…

DGX agent

VCs are now sharing screenshots in group chats of Claude discouraging investment in open-source AI infra startups and models. Obviously there is an absolute EXPLOSION of pitches in inference companies

model-releasesclem-delangue--x
27 Jun 2026
Model Releases

Know2Guess: A Contamination-Aware Multi-Zone Benchmark for Knowledge-Boundary Evaluation in Large Language Models

DGX agent

arXiv:2606.26101v1 Announce Type: cross Abstract: Reliable evaluation of large language models should separate supported answering from unsupported guessing without conflating either with data contami

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

DGX agent

arXiv:2606.26968v1 Announce Type: new Abstract: Speech-capable models are increasingly deployed in real-world applications across languages. Yet their safety and fairness beyond English settings and u

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

DGX agent

arXiv:2510.09976v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and pi_0 have shown strong generalization by leveraging large-scale demonstrations, yet t

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

DGX agent

arXiv:2606.18430v2 Announce Type: replace Abstract: Statistical watermarks help organizations attribute large language model (LLM) outputs, yet existing detectors often struggle when watermark signals

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Are Tabular Foundation Models Robust to Realistic Query Distribution Shifts in Microbiome Data?

DGX agent

arXiv:2606.24995v1 Announce Type: new Abstract: Tabular foundation models (TFMs) achieve strong performance on microbiome abundance data, yet their robustness under realistic distribution shift remain

model-releasesarxiv-cs-lg
25 Jun 2026
Research

Certification of Machine Learning Models via Directional Sharpness

DGX agent

arXiv:2606.25004v1 Announce Type: new Abstract: In machine learning, model certification has been identified as an important method for gaining assurance about a model's trustworthiness and quality. A

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Efficient Remote Sensing Instance Segmentation with Linear-Time State Space Distilled Visual Foundation Models

DGX agent

arXiv:2606.25324v1 Announce Type: new Abstract: The computational complexity of Transformers scales quadratically with the number of tokens, which significantly constrains the efficiency of vision mod

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

EPTS: Elastic Post-Training Sparsity for Efficient Large Language Model Compression

DGX agent

arXiv:2606.25285v1 Announce Type: new Abstract: Post-Training Sparsity (PTS) has emerged as a crucial paradigm for compressing Large Language Models to facilitate efficient deployment on resource-cons

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Perfect Detection, Failed Control: The Geometry of Knowing vs. Steering in Language Models

DGX agent

arXiv:2606.24952v1 Announce Type: new Abstract: A central aspiration of mechanistic interpretability is controllability: if we know where a behavior is represented in a model's activations, we should

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

Quantization Inflates Reasoning: Token Inflation as a Hidden Cost of Low-Bit Reasoning Models

DGX agent

arXiv:2606.25519v1 Announce Type: cross Abstract: Quantization is widely used to reduce the inference cost of large language models, but its effect on reasoning models is not fully captured by final-a

agentsarxiv-cs-lg
25 Jun 2026
Model Releases

Robustness assessment of large audio language models in multiple-choice evaluation

DGX agent

arXiv:2510.04584v2 Announce Type: replace Abstract: Recent advances in large audio language models (LALMs) have primarily been assessed using a multiple-choice question answering (MCQA) framework. How

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

RoDyn: Taming Interactive Robot-Dynamic 2.5D World Model for Robotic Manipulation

DGX agent

arXiv:2510.09036v2 Announce Type: replace Abstract: Learned world models hold significant potential as neural simulators for robotic manipulation. However, prevalent 2D video-based models inherently l

agentsarxiv-cs-ro
25 Jun 2026
Model Releases

Same Evidence, Different Answer: Auditing Order Sensitivity in Multimodal Large Language Models

DGX agent

arXiv:2606.26079v1 Announce Type: new Abstract: Standard benchmarks for multimodal large language models (MLLMs) score each item on one canonical ordering and miss whether order-irrelevant shuffling c

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Small Initialization Matters for Large Language Models

DGX agent

arXiv:2606.17945v2 Announce Type: replace Abstract: Large language models provide a tractable system for asking how intelligence itself emerges, rather than only how LLMs can be engineered. Although p

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models

DGX agent

arXiv:2606.24962v1 Announce Type: new Abstract: Recent progress in large-scale sequence modeling has shown that a single model can learn useful representations across highly diverse data distributions

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability

DGX agent

arXiv:2606.24589v1 Announce Type: new Abstract: Scaling adversarial evaluation of large language models requires both a method for generating hard inputs and a reliable way to confirm that resulting f

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Business as Rulesual: A Benchmark and Framework for Business Rule Flow Modeling with LLMs

DGX agent

arXiv:2505.18542v4 Announce Type: replace Abstract: Extracting structured procedural knowledge from unstructured business documents is a critical yet unresolved bottleneck in process automation. While

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Can Scale Save Us From Plasticity Loss in Large Language Models?

DGX agent

arXiv:2606.24752v1 Announce Type: new Abstract: The loss of plasticity - the ability of a network to learn new information after having already learned older information - is a fundamental challenge i

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Extended pseudo-spectral physics-informed neural networks for phase-field models

DGX agent

arXiv:2606.24660v1 Announce Type: cross Abstract: Phase-field models play a central role in the continuum description of phase separation, in which the bulk free-energy density and the interfacial thi

model-releasesarxiv-cs-lg
24 Jun 2026
Applications

HyMaTE: A Hybrid Mamba and Transformer Model for EHR Representation Learning

DGX agent

arXiv:2509.24118v2 Announce Type: replace Abstract: Electronic health Records (EHRs) have become a cornerstone in modern-day healthcare. They are a crucial part for analyzing the progression of patien

applicationsarxiv-cs-lg
24 Jun 2026
Model Releases

Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients

DGX agent

arXiv:2606.24774v1 Announce Type: new Abstract: Vision-Language Large Models (VLLMs) trained on massive crawled corpora raise pressing copyright and data-provenance concerns. These concerns are partic

model-releasesarxiv-cs-cv
24 Jun 2026
Research

Sentence-Level Contextual Entrainment in Large Language Models

DGX agent

arXiv:2606.24077v1 Announce Type: new Abstract: Contextual entrainment, which is a newly discovered phenomenon in large language models (LLMs), refers to the tendency of a model to assign higher proba

researcharxiv-cs-cl
24 Jun 2026
Agents

The next scaling law is multi-agent swarms. Mixture of models.

DGX agent

The next scaling law is multi-agent swarms. Mixture of models. Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API. Our ‘Fugu Ultra’ model matches the pe

agentsdavid-ha--x
24 Jun 2026
Model Releases

Transformer-Based Language Models Across Domain Verticals: Architectures, Applications and Critical Assessment

DGX agent

arXiv:2606.24331v1 Announce Type: new Abstract: Transformer-based language models have become the default substrate for natural language processing and the pace of new releases has made it hard for pr

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

A-Evolve-Training: Autonomous Post-Training of a 30B Model

DGX agent

arXiv:2606.20657v1 Announce Type: cross Abstract: Post-training a frontier model is normally weeks of human work: proposing data and recipe changes, launching runs, reading evals, deciding what to kee

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

A Linear Fractional Transformation Model and Calibration Method for Light Field Camera

DGX agent

arXiv:2511.03962v2 Announce Type: replace Abstract: Accurate intrinsic calibration is a crucial yet challenging prerequisite for 3D reconstruction using light field cameras. Existing calibration model

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Attacking the Trusted Imagination: Oracle-Level Integrity Attacks on Imagine-then-Act World Models

DGX agent

arXiv:2606.22966v1 Announce Type: new Abstract: Many recent vision-language-action (VLA) policies adopt an imagine-then-act design. A world-action model (WAM) first imagines a short future as a latent

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models

DGX agent

arXiv:2606.21645v1 Announce Type: cross Abstract: Large language models (LLMs) can readily reproduce conventional expressions, yet their ability to model gradient frequency distributions remains under

safetyarxiv-cs-lg
23 Jun 2026
Tutorials

Diffusion-Driven State Space Models

DGX agent

arXiv:2606.21036v1 Announce Type: cross Abstract: In many domains, practitioners seek models that produce accurate forecasts while faithfully capturing latent system dynamics. Existing approaches typi

tutorialsarxiv-cs-lg
23 Jun 2026
Model Releases

Evaluating Large Language Models for Hausa and Fongbe Machine Translation: Benchmarks, Failures, and Metric Reliability

DGX agent

arXiv:2606.22269v1 Announce Type: cross Abstract: We investigate the translation quality of current large language models (LLMs) for English-to-Hausa and English-to-Fongbe - two typologically distinct

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

How Well Do Self-Supervised Speech Models Encode Age and Gender in Children's Speech? A Layer-Wise Analysis Across Multiple Architectures

DGX agent

arXiv:2606.22177v1 Announce Type: cross Abstract: Self-supervised learning (SSL) models have become a central component of modern speech processing systems, as they enable the learning of rich acousti

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

L20-Edu-135M: An Auditable Single-GPU Study of Data-Efficient Small Language Modeling

DGX agent

arXiv:2606.22189v1 Announce Type: new Abstract: Small language models are cheap to serve and feasible on local hardware, but strong public 135M-class systems are commonly trained with hundreds of bill

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Learning-Based Modeling of Soft Robots via Cosserat Rod Theory

DGX agent

arXiv:2606.20958v1 Announce Type: new Abstract: Modeling soft robot dynamics is challenging due to their continuum structure and typically nonlinear dynamics. Creating models based on first-order prin

researcharxiv-cs-ro
23 Jun 2026
Model Releases

LIBERO-Safety: A Comprehensive Benchmark for Physical and Semantic Safety in Vision-Language-Action Models

DGX agent

arXiv:2606.23686v1 Announce Type: new Abstract: Despite the impressive manipulation capabilities of Vision-Language-Action (VLA) models, their operational safety under strict constraints remains large

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Load Testing for Machine Learning Model Serving Systems at Scale

DGX agent

arXiv:2606.22013v1 Announce Type: new Abstract: Machine learning (ML) model serving has become a dominant consumer of GPU infrastructure, yet capacity planning in these systems remains largely ad hoc.

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models

DGX agent

arXiv:2509.23729v3 Announce Type: replace Abstract: Large Language Models (LLMs) with multimodal capabilities have revolutionized vision-language tasks, but their deployment often requires huge memory

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Model Inversion meets Cryptographic Fuzzy Extractors

DGX agent

arXiv:2510.25687v3 Announce Type: replace-cross Abstract: Model inversion attacks pose an open challenge to privacy-sensitive applications that use machine learning (ML) models. For example, face auth

researcharxiv-cs-lg
23 Jun 2026
Safety

PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

DGX agent

arXiv:2606.18375v2 Announce Type: replace Abstract: World foundation models (WFMs) are powerful simulators, yet they predominantly operate in a single-view setting and lack the multi-view 3D consisten

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

Post-Training Speech Enhancement Language Models with Perceptual Rewards

DGX agent

arXiv:2606.21458v1 Announce Type: new Abstract: Speech enhancement language models achieve strong results when trained on discrete audio tokens, but their optimization relies on token-level cross-entr

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Residue-Level Attributions in Protein Language Models Do Not Recover Allergen Epitopes

DGX agent

arXiv:2606.22181v1 Announce Type: new Abstract: Deep allergenicity classifiers are increasingly used in safety screening of novel foods, and recent protein language models have substantially improved

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Translating Inference-Time Control to Radiology Vision-Language Models: Activation Steering for Pneumonia Classification on Chest X-rays

DGX agent

arXiv:2606.20852v1 Announce Type: new Abstract: Inference-time engineering can alter model behavior without fine-tuning. However, its utility for improving diagnostic performance in medical vision-lan

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training

DGX agent

arXiv:2508.03058v2 Announce Type: replace Abstract: Reinforcement Learning (RL) in real-world environments often suffers from ambiguous or incomplete reward supervision, which undermines policy stabil

safetyarxiv-cs-lg
23 Jun 2026
← Previous
1…8081828384…1261
Next →