AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
11 Jun 2026

Tac-DINO: Learning Vision-Tactile Features with Patch Alignment

Model ReleasesDGX agent

arXiv:2606.12069v1 Announce Type: new Abstract: Touch is the primary medium through which humans interact with the environment. Currently, tactile learning mainly focuses on image-level pretraining or

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

Model ReleasesDGX agent

arXiv:2606.12387v1 Announce Type: cross Abstract: Large Language Models (LLMs) have democratized database access through Text-to-SQL, but moving from prototypes to production remains difficult. Real d

Task-Aligned Stability Analysis of Vision-Language Models for Autonomous Driving Hazard Detection

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.11889v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used for scene understanding in autonomous driving, but robustness analysis often relies on task-agnost

Teaching Diffusion to Speculate Left-to-Right

Model ReleasesDGX agent

arXiv:2606.11552v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their autoregressive decoding process incurs substantial i

The Language You Ask In: Language-Conditioned Ideological Divergence in LLM Analysis of Contested Political Documents

Model ReleasesDGX agent

arXiv:2601.12164v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as analytical tools across multilingual contexts, yet their outputs may carry systemati

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of…

Model ReleasesDGX agent

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of code that’s not well understood that you will have to pay l

The N-Body Problem: Parallel Execution from Single-Person Egocentric Video

Model ReleasesDGX agent

arXiv:2512.11393v2 Announce Type: replace Abstract: Humans can intuitively parallelise complex activities, but can a model predict this from observing a single person? Given one egocentric video, we i

The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics

Model ReleasesDGX agent

arXiv:2606.12289v1 Announce Type: cross Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling

The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content

Model ReleasesDGX agent

arXiv:2606.11198v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems inject external knowledge to improve LLM outputs, yet the format of injected content -- distinct from its

Time-multiplexed layer reuse for physical neural networks

Model ReleasesDGX agent

arXiv:2511.00044v3 Announce Type: replace Abstract: Physical neural networks (PNNs) are promising candidates for next-generation computing, but existing demonstrations remain several orders of magnitu

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

Model ReleasesDGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

TimeRouter: Efficient and Adaptive Routing of Time-Series Foundation Models

Model ReleasesDGX agent

arXiv:2606.11625v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) are increasingly explored as predictive experts within emerging agentic time-series systems. However, TSFMs exhibi

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

Model ReleasesDGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

Model ReleasesDGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

Toward Trustworthy AI: Multi-Target Adversarial Attacks and Robust Defenses for Continuous Data Summarization

Model ReleasesDGX agent

arXiv:2606.11804v1 Announce Type: new Abstract: Trustworthy AI requires reliable data-processing pipelines, not only robust downstream predictive models. As an upstream component, data summarization d

Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering

Model ReleasesDGX agent

arXiv:2606.11836v1 Announce Type: cross Abstract: This paper presents a novel data-free and training-free compression approach for speech foundation models using channelwise clustering via k-means. Mo

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

Model ReleasesDGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fabl…

Model ReleasesDGX agent

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fable, this may no longer be possible: - One of our team members

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

Model ReleasesDGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

Model ReleasesDGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

VICX: Generalizable Robot Manipulation via Video Generation and In-Context Operator Network

Model ReleasesDGX agent

arXiv:2606.12028v1 Announce Type: new Abstract: Generalizable robot manipulation requires not only task-level reasoning over unseen scenes, but also reliable grounding of visual plans into embodiment-

VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation

Model ReleasesDGX agent

arXiv:2601.03792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable proficiency in general medical domains. However, their performance significantly degrades

Visualizing LLM Latent Space Geometry Through Dimensionality Reduction

Model ReleasesDGX agent

arXiv:2511.21594v3 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art results across many natural language tasks, but their internal mechanisms remain difficult to

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

Model ReleasesDGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.11906v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in language-conditioned robotic manipulation, yet their robustness to linguistic varia

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications

Model ReleasesDGX agent

arXiv:2601.22025v2 Announce Type: replace-cross Abstract: Evaluating Large Language Model (LLM) applications differs from conventional software testing because outputs are probabilistic, semantically

When Roleplaying, Do Models Believe What They Say?

Model ReleasesDGX agent

arXiv:2606.11502v1 Announce Type: cross Abstract: Language models can state that 'the Earth orbits the Sun' and, when role-playing Aristotle, assert the opposite. Recent work argues that persona adopt

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs

Model ReleasesDGX agent

arXiv:2606.12385v1 Announce Type: new Abstract: Modern LLM training pipelines increasingly rely on other models to generate data, filter corpora, judge outputs, and guide development decisions. These

Wordle 1,817 3/6 ⬛⬛🟨⬛⬛ 🟨🟨🟨🟩⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle puzzle solution (puzzle #1,817) completed in 3 attempts out of 6 allowed guesses, showing the progression of letter feedback (gray for incorrect letters, yellow for correc

World Model Self-Distillation: Training World Models to Solve General Tasks

Model ReleasesDGX agent

arXiv:2606.12072v1 Announce Type: new Abstract: Pretrained video generators are promising visual world models that exhibit emergent task-solving abilities; however, their reliance on detailed textual

World Pilot: Steering Vision-Language-Action Models with World-Action Priors

Model ReleasesDGX agent

arXiv:2606.12403v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit semantic grounding from large-scale pretraining and perform competently across in-distribution manipulation

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

Model ReleasesDGX agent

arXiv:2606.11816v1 Announce Type: cross Abstract: Forecasting real-world events requires language-model agents to reason under uncertainty from incomplete, time-bounded information. Yet evaluating whe

10 Jun 2026

5% > 100%: Flatness Preference is All You Need for Multimodal Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10488v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods provide a streamlined and efficient tool for adapting large models to domain-specific multimodal downstre

A complementary study on PlanGPT: Evaluation with defined Performance Metrics and comparison with a planner

Model ReleasesDGX agent

arXiv:2606.10489v1 Announce Type: new Abstract: Automated Planning is a subfield of Artificial Intelligence (AI) where the main objective is generating a sequence of actions, known as a plan, that hel

A Constrained Natural-Language Interface for Variational Multi-Physics Finite Element Simulations in FEniCS

Model ReleasesDGX agent

arXiv:2606.10928v1 Announce Type: cross Abstract: Large language models can reduce the manual effort required to set up finite element simulations, but they introduce reliability risks when generated

A History-Aware Visually Grounded Critic for Computer Use Agents

Model ReleasesDGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

A Large Scale Open-Source Image and Video Dataset for Robust Wildfire Detection and Classification

Model ReleasesDGX agent

arXiv:2606.10174v1 Announce Type: new Abstract: Wildfire detection and monitoring are critical for mitigating fire spread and reducing environmental and infrastructural damage. In this work, we introd

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI

Model ReleasesDGX agent

arXiv:2505.01458v2 Announce Type: replace-cross Abstract: Navigation and manipulation are core capabilities in Embodied AI, but training agents to perform them directly in the real world is costly, ti

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

Model ReleasesDGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design

Model ReleasesDGX agent

arXiv:2606.10493v1 Announce Type: cross Abstract: Local deployment of large Mixture-of-Experts (MoE) models falls short of the service quality achieved in cloud-scale environments, even under low-conc

AdaGC: Enhancing LLM Pretraining Stability via Adaptive Gradient Clipping

Model ReleasesDGX agent

arXiv:2502.11034v3 Announce Type: replace Abstract: Loss spikes remain a persistent obstacle in large-scale language model pretraining. While previous research has attempted to identify the root cause

Advancing the State-of-the-Art in Empirical Privacy Auditing

Model ReleasesDGX agent

arXiv:2606.10481v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning of large language models (LLMs) can exhibit problematic memorization of individual training examples. Empirical privac

Agentic Hybrid RAG for Evidence-Grounded Muon Collider Analysis

Model ReleasesDGX agent

arXiv:2606.10381v1 Announce Type: cross Abstract: Muon collider research spans accelerator physics, detector instrumentation, and high-energy phenomenology, with relevant evidence scattered across a r

AgenticNav: Zero-Shot Vision-and-Language Navigation as a Tool-Calling Harness

Model ReleasesDGX agent

arXiv:2606.10577v1 Announce Type: new Abstract: Zero-shot vision-and-language navigation in continuous environments (VLN-CE) has recently become feasible with large vision-language models (VLMs). Howe

AgniNav: Configuration-Driven Cross-Embodiment Local Planning for Robot Navigation

Model ReleasesDGX agent

arXiv:2606.10903v1 Announce Type: new Abstract: Monocular local navigation is attractive for lightweight robots, but existing vision-based policies often couple perception to a specific body, camera h

[AINews] Anthropic Claude Fable 5 — Mythos but Safe, with Controversial Terms

Model ReleasesDGX agent

This article from Latent Space discusses Anthropic's Claude Fable 5 model, examining its capabilities in handling creative and mythological content while maintaining safety guardrails, along with cove

Alignment Collapse Under KV Cache Quantization: Diagnosis and Mitigation

Model ReleasesDGX agent

arXiv:2606.09864v1 Announce Type: cross Abstract: Key-value (KV) cache quantization is widely used to reduce Large Language Model (LLM) inference memory, yet existing evaluations solely focus on measu

An adaptive framework for the axisymmetric pulsar magnetosphere using physics-informed Kolmogorov-Arnold networks

Model ReleasesDGX agent

arXiv:2606.10686v1 Announce Type: cross Abstract: The pulsar magnetosphere has only recently been addressed using Physics-Informed Neural Networks (PINNs), by deploying a domain-decomposition approach

An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs

Model ReleasesDGX agent

arXiv:2603.14463v2 Announce Type: replace Abstract: Adapting Large Language Models (LLMs) to high-stakes vertical domains like insurance presents a significant challenge: scenarios demand strict adher

and the video for reference: https://x.com/ClaudeDevs/status/2064399512664526853 (I didnt get to use the updated designs in time)

Model ReleasesDGX agent

and the video for reference: https://x.com/ClaudeDevs/status/2064399512664526853 (I didnt get to use the updated designs in time) Claude Fable 5 changed how we work on the Claude Code team day to day.

Announcing the Gemma challenge! Google, Hugging Face, and the open-source AI community choose to empower AI builders rather than sabotage th…

Model ReleasesDGX agent

Announcing the Gemma challenge! Google, Hugging Face, and the open-source AI community choose to empower AI builders rather than sabotage them. Fun to see the Hub becoming the platform where agents co

Anomaly Detection and Root Cause Analysis for Microservice Systems

Model ReleasesDGX agent

arXiv:2606.09942v1 Announce Type: cross Abstract: Microservice systems are widely used to build cloud applications, yet their complexity makes failures inevitable, degrading user experience and causin

Anthropic backtracks on a policy limiting Claude Fable 5's ability to develop other AI models, after significant backlash from the AI research community (Maxwell Zeff/Wired)

Model ReleasesDGX agent

Maxwell Zeff / Wired: Anthropic backtracks on a policy limiting Claude Fable 5's ability to develop other AI models, after significant backlash from the AI research community — The company changed cou

Anthropic secretly limiting Claude's usefulness for LLM development strengthens the argument that Anthropic is using AI safety to justify monopolistic behavior (Dean W. Ball/@deanwball)

Model ReleasesDGX agent

Dean W. Ball / @deanwball: Anthropic secretly limiting Claude's usefulness for LLM development strengthens the argument that Anthropic is using AI safety to justify monopolistic behavior — My last obs

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to buil…

Model ReleasesDGX agent

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to building pretraining pipelines, distributed training infrastruct

ASA: Backbone-Training-Free Representation Engineering for Tool-Calling Agents

Model ReleasesDGX agent

arXiv:2602.04935v3 Announce Type: replace-cross Abstract: Adapting LLM agents to domain-specific tool calling remains notably brittle under evolving interfaces. Prompt and schema engineering is easy t

Assessing Automated Prompt Injection Attacks in Agentic Environments

Model ReleasesDGX agent

arXiv:2606.10525v1 Announce Type: cross Abstract: Indirect prompt injection poses a critical threat to LLM agents that interact with untrusted external data, yet automated attack methods--proven effec

ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark

Model ReleasesDGX agent

arXiv:2505.23851v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied to symbolic mathematics, yet existing evaluations often conflate pattern memorization wi

at Code w/ Claude Tokyo! say hi if you see me around

Model ReleasesDGX agent

Thariq posted about attending a coding event or workshop featuring Claude in Tokyo and invited others to say hello if they encountered him there. The post appears to be a social announcement about his

Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It

Model ReleasesDGX agent

arXiv:2606.11052v1 Announce Type: new Abstract: Chain-of-thought (CoT) supervised fine-tuning (SFT) is widely adopted to improve reasoning ability, yet we find that it systematically degrades long-con

← Previous
1…148149150151152…377
Next →