AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
All
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

Teaching Diffusion to Speculate Left-to-Right

DGX agent

arXiv:2606.11552v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their autoregressive decoding process incurs substantial i

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

The Language You Ask In: Language-Conditioned Ideological Divergence in LLM Analysis of Contested Political Documents

DGX agent

arXiv:2601.12164v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as analytical tools across multilingual contexts, yet their outputs may carry systemati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of…

DGX agent

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of code that’s not well understood that you will have to pay l

model-releasesjerry-liu--x
11 Jun 2026
Model Releases

The N-Body Problem: Parallel Execution from Single-Person Egocentric Video

DGX agent

arXiv:2512.11393v2 Announce Type: replace Abstract: Humans can intuitively parallelise complex activities, but can a model predict this from observing a single person? Given one egocentric video, we i

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics

DGX agent

arXiv:2606.12289v1 Announce Type: cross Abstract: As Artificial Intelligence models grow in complexity, interpretability has become an indispensable tool for understanding, debugging, and controlling

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content

DGX agent

arXiv:2606.11198v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems inject external knowledge to improve LLM outputs, yet the format of injected content -- distinct from its

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Time-multiplexed layer reuse for physical neural networks

DGX agent

arXiv:2511.00044v3 Announce Type: replace Abstract: Physical neural networks (PNNs) are promising candidates for next-generation computing, but existing demonstrations remain several orders of magnitu

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation

DGX agent

arXiv:2606.11990v1 Announce Type: cross Abstract: Remaining Useful Life (RUL) prediction is essential for industrial predictive maintenance, yet many learning-based approaches rely on extensive featur

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

TimeRouter: Efficient and Adaptive Routing of Time-Series Foundation Models

DGX agent

arXiv:2606.11625v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) are increasingly explored as predictive experts within emerging agentic time-series systems. However, TSFMs exhibi

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

DGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

DGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Toward Trustworthy AI: Multi-Target Adversarial Attacks and Robust Defenses for Continuous Data Summarization

DGX agent

arXiv:2606.11804v1 Announce Type: new Abstract: Trustworthy AI requires reliable data-processing pipelines, not only robust downstream predictive models. As an upstream component, data summarization d

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering

DGX agent

arXiv:2606.11836v1 Announce Type: cross Abstract: This paper presents a novel data-free and training-free compression approach for speech foundation models using channelwise clustering via k-means. Mo

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

DGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fabl…

DGX agent

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fable, this may no longer be possible: - One of our team members

model-releasesjerry-liu--x
11 Jun 2026
Model Releases

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

DGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

DGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

VICX: Generalizable Robot Manipulation via Video Generation and In-Context Operator Network

DGX agent

arXiv:2606.12028v1 Announce Type: new Abstract: Generalizable robot manipulation requires not only task-level reasoning over unseen scenes, but also reliable grounding of visual plans into embodiment-

model-releasesarxiv-cs-ro
11 Jun 2026
Model Releases

VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation

DGX agent

arXiv:2601.03792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable proficiency in general medical domains. However, their performance significantly degrades

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Visualizing LLM Latent Space Geometry Through Dimensionality Reduction

DGX agent

arXiv:2511.21594v3 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art results across many natural language tasks, but their internal mechanisms remain difficult to

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

DGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

DGX agent

arXiv:2606.11906v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in language-conditioned robotic manipulation, yet their robustness to linguistic varia

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications

DGX agent

arXiv:2601.22025v2 Announce Type: replace-cross Abstract: Evaluating Large Language Model (LLM) applications differs from conventional software testing because outputs are probabilistic, semantically

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

When Roleplaying, Do Models Believe What They Say?

DGX agent

arXiv:2606.11502v1 Announce Type: cross Abstract: Language models can state that 'the Earth orbits the Sun' and, when role-playing Aristotle, assert the opposite. Recent work argues that persona adopt

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs

DGX agent

arXiv:2606.12385v1 Announce Type: new Abstract: Modern LLM training pipelines increasingly rely on other models to generate data, filter corpora, judge outputs, and guide development decisions. These

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Wordle 1,817 3/6 ⬛⬛🟨⬛⬛ 🟨🟨🟨🟩⬛ 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle puzzle solution (puzzle #1,817) completed in 3 attempts out of 6 allowed guesses, showing the progression of letter feedback (gray for incorrect letters, yellow for correc

model-releasesanthropic--x
11 Jun 2026
Model Releases

World Model Self-Distillation: Training World Models to Solve General Tasks

DGX agent

arXiv:2606.12072v1 Announce Type: new Abstract: Pretrained video generators are promising visual world models that exhibit emergent task-solving abilities; however, their reliance on detailed textual

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

World Pilot: Steering Vision-Language-Action Models with World-Action Priors

DGX agent

arXiv:2606.12403v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit semantic grounding from large-scale pretraining and perform competently across in-distribution manipulation

model-releasesarxiv-cs-ro
11 Jun 2026
Model Releases

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

DGX agent

arXiv:2606.11816v1 Announce Type: cross Abstract: Forecasting real-world events requires language-model agents to reason under uncertainty from incomplete, time-bounded information. Yet evaluating whe

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

5% > 100%: Flatness Preference is All You Need for Multimodal Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2606.10488v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods provide a streamlined and efficient tool for adapting large models to domain-specific multimodal downstre

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

A complementary study on PlanGPT: Evaluation with defined Performance Metrics and comparison with a planner

DGX agent

arXiv:2606.10489v1 Announce Type: new Abstract: Automated Planning is a subfield of Artificial Intelligence (AI) where the main objective is generating a sequence of actions, known as a plan, that hel

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

A Constrained Natural-Language Interface for Variational Multi-Physics Finite Element Simulations in FEniCS

DGX agent

arXiv:2606.10928v1 Announce Type: cross Abstract: Large language models can reduce the manual effort required to set up finite element simulations, but they introduce reliability risks when generated

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

A History-Aware Visually Grounded Critic for Computer Use Agents

DGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

A Large Scale Open-Source Image and Video Dataset for Robust Wildfire Detection and Classification

DGX agent

arXiv:2606.10174v1 Announce Type: new Abstract: Wildfire detection and monitoring are critical for mitigating fire spread and reducing environmental and infrastructural damage. In this work, we introd

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI

DGX agent

arXiv:2505.01458v2 Announce Type: replace-cross Abstract: Navigation and manipulation are core capabilities in Embodied AI, but training agents to perform them directly in the real world is costly, ti

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

DGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design

DGX agent

arXiv:2606.10493v1 Announce Type: cross Abstract: Local deployment of large Mixture-of-Experts (MoE) models falls short of the service quality achieved in cloud-scale environments, even under low-conc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

AdaGC: Enhancing LLM Pretraining Stability via Adaptive Gradient Clipping

DGX agent

arXiv:2502.11034v3 Announce Type: replace Abstract: Loss spikes remain a persistent obstacle in large-scale language model pretraining. While previous research has attempted to identify the root cause

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Advancing the State-of-the-Art in Empirical Privacy Auditing

DGX agent

arXiv:2606.10481v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning of large language models (LLMs) can exhibit problematic memorization of individual training examples. Empirical privac

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Agentic Hybrid RAG for Evidence-Grounded Muon Collider Analysis

DGX agent

arXiv:2606.10381v1 Announce Type: cross Abstract: Muon collider research spans accelerator physics, detector instrumentation, and high-energy phenomenology, with relevant evidence scattered across a r

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

AgenticNav: Zero-Shot Vision-and-Language Navigation as a Tool-Calling Harness

DGX agent

arXiv:2606.10577v1 Announce Type: new Abstract: Zero-shot vision-and-language navigation in continuous environments (VLN-CE) has recently become feasible with large vision-language models (VLMs). Howe

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

AgniNav: Configuration-Driven Cross-Embodiment Local Planning for Robot Navigation

DGX agent

arXiv:2606.10903v1 Announce Type: new Abstract: Monocular local navigation is attractive for lightweight robots, but existing vision-based policies often couple perception to a specific body, camera h

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

[AINews] Anthropic Claude Fable 5 — Mythos but Safe, with Controversial Terms

DGX agent

This article from Latent Space discusses Anthropic's Claude Fable 5 model, examining its capabilities in handling creative and mythological content while maintaining safety guardrails, along with cove

model-releaseslatent-space
10 Jun 2026
Model Releases

Alignment Collapse Under KV Cache Quantization: Diagnosis and Mitigation

DGX agent

arXiv:2606.09864v1 Announce Type: cross Abstract: Key-value (KV) cache quantization is widely used to reduce Large Language Model (LLM) inference memory, yet existing evaluations solely focus on measu

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

An adaptive framework for the axisymmetric pulsar magnetosphere using physics-informed Kolmogorov-Arnold networks

DGX agent

arXiv:2606.10686v1 Announce Type: cross Abstract: The pulsar magnetosphere has only recently been addressed using Physics-Informed Neural Networks (PINNs), by deploying a domain-decomposition approach

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs

DGX agent

arXiv:2603.14463v2 Announce Type: replace Abstract: Adapting Large Language Models (LLMs) to high-stakes vertical domains like insurance presents a significant challenge: scenarios demand strict adher

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

and the video for reference: https://x.com/ClaudeDevs/status/2064399512664526853 (I didnt get to use the updated designs in time)

DGX agent

and the video for reference: https://x.com/ClaudeDevs/status/2064399512664526853 (I didnt get to use the updated designs in time) Claude Fable 5 changed how we work on the Claude Code team day to day.

model-releasesthariq--x
10 Jun 2026
Model Releases

Announcing the Gemma challenge! Google, Hugging Face, and the open-source AI community choose to empower AI builders rather than sabotage th…

DGX agent

Announcing the Gemma challenge! Google, Hugging Face, and the open-source AI community choose to empower AI builders rather than sabotage them. Fun to see the Hub becoming the platform where agents co

model-releasesclem-delangue--x
10 Jun 2026
← Previous
1…187188189190191…472
Next →