AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlog
90,223Total entries
1Added by human
90,222Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,101 results
Model Releases

DLR: Zero-Inference-Cost Latent Residuals for Low-Rank Pre-Training

DGX agent

arXiv:2606.28932v1 Announce Type: cross Abstract: Large language models have driven recent progress in language and multimodal AI, yet pre-training them at scale is prohibitively expensive. Low-rank p

model-releasesarxiv-cs-ai
30 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training

DGX agent

arXiv:2606.30345v1 Announce Type: cross Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Dynamic Parsing and Updating Natural Language Specification using VLMs for Robust Vision-Language Tracking

DGX agent

arXiv:2606.29357v1 Announce Type: cross Abstract: Vision-language tracking guided by natural language specifications leverages high-level semantic cues of target objects to substantially boost trackin

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Dynamo: Dynamic Skill-Tool Evolution for Vision-Language Agents

DGX agent

arXiv:2606.30185v1 Announce Type: new Abstract: Improving vision-language models (VLMs) on visual reasoning typically requires retraining or hand-designed prompts and tools. We present Dynamo, a train

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

EVLA: An Electro-Aware Multimodal Assistant for Physically-Grounded Driving Reasoning and Control

DGX agent

arXiv:2606.28938v1 Announce Type: new Abstract: Modern vision-language models (VLMs) for driving assistants typically treat vehicle dynamics as a black box, resulting in decisions that lack awareness

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Experience Augmented Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.30420v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a powerful paradigm for improving the reasoning capabilities of large language models (LLMs). H

model-releasesarxiv-cs-lg
30 Jun 2026
Research

Few-Step Boltzmann Generators via Scalable Likelihood Flow Maps

DGX agent

arXiv:2606.29110v1 Announce Type: new Abstract: Recent progress in flow-based generative modeling has led to models that output high-quality samples while using only a small number of function evaluat

researcharxiv-cs-lg
30 Jun 2026
Model Releases

Generalization Analysis of Transformers in Distribution Regression

DGX agent

arXiv:2606.29256v1 Announce Type: cross Abstract: In recent years, models based on the Transformer architecture have seen widespread applications and have become one of the core tools in the field of

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Heterogeneous Tactile Transformer

DGX agent

arXiv:2606.29948v1 Announce Type: new Abstract: Tactile sensors are inherently heterogeneous: a model trained on one sensor cannot be directly used on another, which limits learning contact-rich manip

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

How Far Can You Get Without a GPU? A Systematic Benchmark of Lightweight Hallucination Detection Across Question Answering, Dialogue, and Summarisation

DGX agent

arXiv:2606.29809v1 Announce Type: cross Abstract: Hallucination detection has become a pressing requirement for trustworthy AI deployment at scale. The most accurate detection methods depend on GPU-in

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How much of an LLM-generated clinical corpus is actually new? A production-scale measurement of content redundancy for provenance classification

DGX agent

arXiv:2606.29605v1 Announce Type: new Abstract: Clinical machine learning increasingly relies on training corpora generated by large language models (LLMs) rather than annotated by clinicians, and suc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

I-BBS: Coordinate-Free Inference of Latent Sub-Manifolds Using Random Distance Matrix Theory

DGX agent

arXiv:2606.29675v1 Announce Type: new Abstract: Bogomolny, Bohigas and Schmit (BBS) found that the spectrum of the pairwise distance matrix on N points sampled from a smooth d-dimensional manifold enc

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Intermediate Text Representation Guided Text-to-Image Generation for Enhancing One-and-Only Alignment

DGX agent

arXiv:2606.30262v1 Announce Type: new Abstract: Text-to-image (T2I) diffusion models often fail to faithfully render explicit textual descriptions, instead defaulting to strongly learned visual priors

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Introducing GeneBench-Pro

DGX agent

GeneBench-Pro is a comprehensive benchmarking tool or dataset introduced by OpenAI designed to evaluate AI model performance on genomic and gene-related tasks. It likely provides standardized metrics

model-releasesopenai
30 Jun 2026
Model Releases

LoGSAM: Parameter-Efficient Cross-Modal Grounding for MRI Segmentation

DGX agent

arXiv:2603.17576v3 Announce Type: replace Abstract: Precise localization and delineation of brain tumors using magnetic resonance imaging (MRI) are essential for planning therapy and guiding surgical

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

LUMEN: Cost-Transparent Multi-Agent Pipeline for Automated Systematic Review and Meta-Analysis

DGX agent

arXiv:2606.28362v1 Announce Type: cross Abstract: Systematic reviews and meta-analyses (SR/MA) remain the gold standard for evidence synthesis, yet completing one typically requires 67 weeks and subst

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Multimodal Graph RAG for Long-range Visually Rich Document Understanding

DGX agent

arXiv:2606.28780v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are widely applied to visual document understanding. However, comprehending long documents remains an issue b

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Multimodal Mathematical Reasoning with Diverse Solving Perspective

DGX agent

arXiv:2507.02804v2 Announce Type: replace Abstract: Recent progress in large-scale reinforcement learning (RL) has notably enhanced the reasoning capabilities of large language models (LLMs), especial

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Nano Banana 2 Lite

DGX agent

Nano Banana 2 Lite Also known as Gemini 3.1 Flash Lite Image (gemini-3.1-flash-lite-image in their API), this is the 'fastest and cheapest Gemini image model, engineered for velocity and scale'. I use

model-releasessimon-willison
30 Jun 2026
Model Releases

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

DGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

OP3DSG: Open-Vocabulary Part-Aware 3D Scene Graph Generation for Real-World Environments

DGX agent

arXiv:2606.29786v1 Announce Type: new Abstract: 3D scene graphs (3DSGs) provide a compact and structured abstraction of 3D environments. Although advances in foundation models have enabled open-vocabu

model-releasesarxiv-cs-cv
30 Jun 2026
Research

Perspectives on Latent Factor Indeterminacy and its Implications for Data Representation

DGX agent

arXiv:2606.28854v1 Announce Type: cross Abstract: The common factor analytic model is related to Helmholtz and Boltzmann machines, can be conceived as a linear autoencoder, or can be thought of as a s

researcharxiv-cs-ai
30 Jun 2026
Safety

REAR: Test-time Preference Realignment through Reward Decomposition

DGX agent

arXiv:2606.30339v1 Announce Type: new Abstract: Aligning large language models (LLMs) with diverse user preferences is a critical yet challenging task. While post-training methods can adapt models to

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

REPAIR-Bench: A Benchmark for Robot Error Perception And Interaction Recovery

DGX agent

arXiv:2606.29937v1 Announce Type: new Abstract: Understanding how users perceive and respond to robot failures is essential for building robust and trustworthy robot systems. Prior work, however, (i)

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Reported Confidence in LLMs Tracks Commitment More Than Correctness

DGX agent

arXiv:2606.29490v1 Announce Type: cross Abstract: Confidence is an estimate of the probability that a chosen answer is correct. Verbal confidence reports are widely used as uncertainty measures in lar

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

DGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing

DGX agent

arXiv:2606.29887v1 Announce Type: new Abstract: In real-world applications, guardrails are often expected to identify unsafe user-model interactions according to application-specific safety policies,

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Semantic-Driven Scale and Spatial Selection for Efficient Cross-Modal Alignment in Referring Remote Sensing Image Segmentation

DGX agent

arXiv:2606.30244v1 Announce Type: new Abstract: Referring Remote Sensing Image Segmentation (RRSIS) seeks to localize and segment the target object or region specified by a natural language expression

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Symbolic Mechanistic Data Attribution: Tracing Training Influence to Learned Behavioral Policies

DGX agent

arXiv:2606.29171v1 Announce Type: cross Abstract: While existing data attribution methods can identify which training examples build specific mechanistic circuits, they cannot explain how training dat

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

TF-MoE: Time-Frequency Mixture-of-Experts for Efficient Speech Separation

DGX agent

arXiv:2606.29575v1 Announce Type: cross Abstract: Recent advances in speech separation (SS) have led to compact front-end models with small parameter sizes, yet their high computational cost remains a

model-releasesarxiv-cs-ai
30 Jun 2026
Research

The Fundamental Limits of Valid Transport Map Estimation

DGX agent

arXiv:2606.30574v1 Announce Type: new Abstract: Many modern generative modeling methods, including diffusion models, normalizing flows, and flow matching, estimate transport maps or plans between dist

researcharxiv-cs-lg
30 Jun 2026
Safety

Towards Physical Intuitions for Alignment Dynamics: A Case Study With Randomness Crystallization

DGX agent

arXiv:2606.29933v1 Announce Type: new Abstract: The alignment of language models is typically studied through the lens of capability benchmarks, but the dynamics of how models change during post-train

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

DGX agent

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception

DGX agent

arXiv:2606.30332v1 Announce Type: new Abstract: Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing app

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

DGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

You Only Touch Once: 6-DoF Object Pose Estimation from Single Tactile Contact

DGX agent

arXiv:2606.28899v1 Announce Type: new Abstract: Accurate 6-DoF object pose estimation is fundamental to robotic manipulation, yet vision-based methods often fail under occlusion, poor lighting, and re

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

A Comparison of Fusion Techniques for Multi-Modal Human Activity Recognition on the HARMES Dataset

DGX agent

arXiv:2606.27886v1 Announce Type: new Abstract: Recent advances in Human Activity Recognition (HAR) from wearable sensors have shown that multi-modal deep learning models consistently outperform their

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Accelerating Attention with Basis Decomposition

DGX agent

arXiv:2510.01718v2 Announce Type: replace Abstract: Attention is a core operation in large language models (LLMs). We present BD Attention (BDA), a lossless algorithmic reformulation of attention. BDA

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's t…

DGX agent

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's the largest AI model theft ever attempted. The same week, the

model-releasesemad-mostaque--x
29 Jun 2026
Model Releases

Applicability of memorization indicators for early spotting of overfitting while recalibrating sEMG-decoders on low sample sizes

DGX agent

arXiv:2606.27855v1 Announce Type: cross Abstract: Deep learning models for surface electromyography (sEMG) can benefit substantially from subject-specific (re-)calibration, since no sufficiently large

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Claude in Microsoft Foundry is now generally available, hosted on Azure. Azure customers get Claude Opus 4.8 and Claude Haiku 4.5, with Azur…

DGX agent

Claude models are now generally available through Microsoft Foundry on Azure infrastructure, providing Azure customers access to Claude Opus 4.8 and Claude Haiku 4.5. This integration allows enterpris

model-releasesboris-cherny--x
29 Jun 2026
Model Releases

Cloud CISO Perspectives: How Google Cloud Security uses AI internally

DGX agent

Welcome to the second Cloud CISO Perspectives for June 2026. Today, we’re discussing how we use AI to chart a path to autonomous software development lifecycle security.As with all Cloud CISO Perspect

model-releasesgoogle-cloud-ai
29 Jun 2026
Model Releases

Contagion Networks: Evaluator Preference Propagation in Multi-Agent LLM Systems

DGX agent

arXiv:2606.20493v2 Announce Type: replace-cross Abstract: When large language models serve as evaluators in multi-agent systems, their strategy preferences -- whether induced by explicit prompts or by

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

Deployment-Side Adaptiveness in Multi-Horizon Volatility Forecasting

DGX agent

arXiv:2606.27688v1 Announce Type: cross Abstract: In financial forecasting, predictive performance depends not only on which model is trained, but also on how the trained model is deployed. We study t

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

Estimation--Prediction Tradeoff in Causal Probabilistic Temporal Graphs

DGX agent

arXiv:2606.28225v1 Announce Type: new Abstract: Temporal link prediction is usually evaluated by predictive performance on unseen edges, but in probabilistic temporal graphs this criterion can conflat

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning

DGX agent

arXiv:2603.09731v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are increasingly considered as a foundation for embodied agents, yet it remains unclear whether they

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

DGX agent

arXiv:2606.27378v1 Announce Type: new Abstract: We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchma

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

From Detection to Action: Using LLM Agents for Fault-Tolerant Control

DGX agent

arXiv:2606.28011v1 Announce Type: cross Abstract: We propose an agentic Large Language Model (LLM) framework for active Fault-Tolerant Control (FTC) that transforms fault detection outputs into constr

model-releasesarxiv-cs-lg
29 Jun 2026
← Previous
1…444445446447448…1357
Next →