AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,813 results
Model Releases

Hallucination Is Linearly Decodable from Mid-Layer Hidden States in Quantized LLMs

DGX agent

arXiv:2606.02628v1 Announce Type: cross Abstract: We investigate whether open-source LLMs encode a linearly separable truthfulness signal in their hidden states, and at which network depth this signal

model-releasesarxiv-cs-cl
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Hallucinations as Orthogonal Noise: Inference-Time Manifold Alignment via Dynamic Contextual Orthogonalization

DGX agent

arXiv:2606.03022v1 Announce Type: cross Abstract: Hallucination in Large Language Models (LLMs), characterized by the generation of content inconsistent with contextual facts or logical constraints --

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Hedge-Bench: Benchmarking Agents on Hard, Realistic Tasks Pertaining to Financial Reasoning

DGX agent

arXiv:2606.03918v1 Announce Type: new Abstract: AI agents can increasingly handle the mechanical tasks of financial analysis: retrieving documents, calculating formulas, updating spreadsheets. The har

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Honesty in Causal Forests: When It Helps and When It Hurts

DGX agent

arXiv:2506.13107v4 Announce Type: replace Abstract: Causal forests estimate how treatment effects vary across individuals, guiding personalized interventions in areas like marketing, operations, and p

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

how does the brain build and track an internal state of the world from (possibly incomplete and noisy) visual observations? i believe visual…

DGX agent

how does the brain build and track an internal state of the world from (possibly incomplete and noisy) visual observations? i believe visual state tracking will be the grand challenge for vision in th

model-releasesyann-lecun--x
3 Jun 2026
Model Releases

How Many Trees in a Random Forest? A Revisited Approach with Plateau Search and Optuna Integration

DGX agent

arXiv:2606.03549v1 Announce Type: new Abstract: Hyperparameter optimization (HPO) for Random Forest faces a specific difficulty in tuning the number of trees: the predictive score typically improves m

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

How Quantization Changes Interpretable Features: A Sparse Autoencoder Analysis of Language Models

DGX agent

arXiv:2606.03002v1 Announce Type: cross Abstract: Quantization is a standard path to deploying large language models, and a quantized model is typically judged acceptable when its perplexity or downst

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

How Wasmer used Codex to build a Node.js runtime for the edge

DGX agent

Wasmer leveraged OpenAI's Codex to develop a Node.js runtime optimized for edge computing environments. The project demonstrates how AI-assisted code generation can accelerate the creation of speciali

model-releasesopenai
3 Jun 2026
Model Releases

Hybrid Dynamics Modeling for a Flexible 2-DoF Robotic Arm

DGX agent

arXiv:2606.02969v1 Announce Type: new Abstract: This paper examines three approaches for modeling the dynamics of a flexible-link 2-DoF robotic arm to address unmodeled dynamics not captured by rigid-

model-releasesarxiv-cs-ro
3 Jun 2026
Model Releases

HyperPatch: Sequential Knowledge Editing Under n-ary Structural Drift

DGX agent

arXiv:2606.03179v1 Announce Type: new Abstract: Large Language Models (LLMs) rely on Knowledge Editing (KE) to maintain temporal validity, yet real-world knowledge is inherently n-ary. We demonstrate

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

I like the racing and Stardew Valley portions, the conclusion is very Claude, though.

DGX agent

This appears to be Ethan Mollick's personal reflection on an AI-generated or AI-assisted project that includes racing and Stardew Valley game elements, with commentary that the conclusion exhibits cha

model-releasesethan-mollick--x
3 Jun 2026
Model Releases

I present to you... The Spaghetti Benchmark

DGX agent

'Will Smith eating spaghetti' became a shorthand for early-stage AI video generation limitations , with the original 2023 clip generated with ModelScope showing distorted faces, morphed hands, and unn

model-releasesr-chatgpt
3 Jun 2026
Model Releases

Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Human and Artificial Cognition

DGX agent

arXiv:2511.21731v2 Announce Type: replace-cross Abstract: We present the results of cognitive tests on conceptual combinations, performed using specific Large Language Models (LLMs) as test subjects.

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

IdiomX A Multilingual Benchmark for Idiom Understanding, Retrieval, and Interpretation

DGX agent

arXiv:2606.02584v1 Announce Type: cross Abstract: Idiomatic expressions remain a persistent challenge for natural language processing because their meanings are often non-compositional, context-depend

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

If this prompt feels well written to you, it's because Suzanne is a writer in her little spare time! You can read her short story, Mall of A…

DGX agent

If this prompt feels well written to you, it's because Suzanne is a writer in her little spare time! You can read her short story, Mall of America here: https://suzannewang.com/mall-of-america It's on

model-releasesthariq--x
3 Jun 2026
Model Releases

Improvise, Adapt, Overcome: An On-The-Fly Multifidelity Algorithm for Efficient Machine Learning

DGX agent

arXiv:2606.02662v1 Announce Type: cross Abstract: Machine learning has accelerated quantum chemistry but is hindered by the prohibitive cost of generating high fidelity training data. Multifidelity ma

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

In early May, the best superforecasters predicted that, by the end of the year, the longest METR 80% task horizons would reach 3-4 hours. In…

DGX agent

In early May, the best superforecasters predicted that, by the end of the year, the longest METR 80% task horizons would reach 3-4 hours. In late May, Claude Mythos achieved that number. We also asked

model-releasesethan-mollick--x
3 Jun 2026
Model Releases

InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning

DGX agent

arXiv:2602.06960v3 Announce Type: replace-cross Abstract: Large reasoning models achieve strong performance by scaling inference-time chain-of-thought, but this paradigm suffers from quadratic cost, c

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Instant Personalized Large Language Model Adaptation via Hypernetwork

DGX agent

arXiv:2510.16282v2 Announce Type: replace Abstract: Personalized large language models (LLMs) tailor content to individual preferences using user profiles or histories. However, existing parameter-eff

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Introducing new capabilities to GPT-Rosalind

DGX agent

GPT-Rosalind is an OpenAI model with newly introduced capabilities designed to enhance its performance in specific tasks or domains. The update likely expands the model's functionality in areas such a

model-releasesopenai
3 Jun 2026
Model Releases

Investigating Adversarial Robustness of Multi-modal Large Language Models

DGX agent

arXiv:2606.03713v1 Announce Type: new Abstract: Multi-modal Large Language Models (MLLMs) achieve strong performance on vision-language tasks, but incorporating visual inputs through a vision encoder

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation

DGX agent

arXiv:2606.03168v1 Announce Type: new Abstract: While instruction-based video editing has seen significant progress, joint audio-visual editing remains constrained by the absence of dedicated datasets

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he h…

DGX agent

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he helped create is over. The agent harness ate the abstraction

model-releasesjerry-liu--x
3 Jun 2026
Model Releases

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

DGX agent

arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coord

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Knowledge Editing in Masked Diffusion Language Models

DGX agent

arXiv:2606.03924v1 Announce Type: new Abstract: Knowledge editing aims to update or correct factual knowledge in a language model. A widely used approach, locate-then-edit, does this in two steps: it

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Knowledge-Preserved Model Tuning in Null-Space for Robust Spatio-Temporal Video Grounding

DGX agent

arXiv:2606.03539v1 Announce Type: new Abstract: Spatio-Temporal Video Grounding aims to localize object tubes based on textual queries. While recent methods have achieved remarkable success, they main

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation

DGX agent

arXiv:2510.22491v3 Announce Type: replace-cross Abstract: Generating high-fidelity 3D geometries under explicit parameter constraints is central to engineering design, yet current methods often requir

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Language Bias under Conflicting Information in Multilingual LLMs

DGX agent

arXiv:2604.07123v2 Announce Type: replace Abstract: Large Language Models (LLMs) have been shown to contain biases in the process of integrating conflicting information when answering questions. Here

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Laplacian Representations for Decision-Time Planning

DGX agent

arXiv:2602.05031v2 Announce Type: replace Abstract: Planning with a learned model remains a key challenge in model-based reinforcement learning (RL). In decision-time planning, state representations a

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

LEAP: Supercharging LLMs for Formal Mathematics with Agentic Frameworks

DGX agent

arXiv:2606.03303v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong informal mathematical reasoning but struggle to generate mechanically verifiable proofs in formal languages

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Learning Temporal Causal Structure via Smooth Differentiable Optimization

DGX agent

arXiv:2606.03227v1 Announce Type: new Abstract: Causal discovery with instantaneous effects in multivariate time series is challenging, as the instantaneous structure must be acyclic. Prior methods en

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution

DGX agent

arXiv:2505.24037v3 Announce Type: replace Abstract: Sparse large language models (LLMs) offer an attractive direction toward efficient deployment, but adapting them to downstream tasks remains challen

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Let the Dynamics Flow: Stable Flow Matching Dynamical Systems

DGX agent

arXiv:2606.03834v1 Announce Type: new Abstract: Flow matching has recently emerged as a powerful approach for imitation learning, enabling scalable, expressive, and multimodal motion policies. However

model-releasesarxiv-cs-ro
3 Jun 2026
Model Releases

LiveBand: Live Accompaniment Generation in the Audio Domain

DGX agent

arXiv:2606.03803v1 Announce Type: cross Abstract: We present LiveBand, a real-time system that generates high-fidelity music accompaniments to live audio input, respecting strict causal constraints. O

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention

DGX agent

arXiv:2606.02680v1 Announce Type: new Abstract: Sparse causal attention is usually described by sequence locality: nearby tokens should remain easy to access, while distant tokens may be dropped to re

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Low-Frequency Shortcuts in Texture-Driven Visual Learning

DGX agent

arXiv:2606.03493v1 Announce Type: new Abstract: Neural networks suffer from shortcut learning, where learned features generalize well to the training set but not to in-distribution (ID) or out-of-dist

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Make sure to update your runtime first! > lms runtime update --all Learn more about this model release https://x.com/googlegemma/status/2062…

DGX agent

Make sure to update your runtime first! > lms runtime update --all Learn more about this model release https://x.com/googlegemma/status/2062202706882883696?s=20 Meet Gemma 4 12B! A unified, encoder-fr

model-releaseslm-studio--x
3 Jun 2026
Model Releases

MARIO: Motion-Augmented Real-Time Multi-Sensor Inertial Odometry

DGX agent

arXiv:2606.02996v1 Announce Type: cross Abstract: Inertial odometry (IO) using only Inertial Measurement Units (IMUs) provides a lightweight solution for human motion tracking in augmented reality (AR

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents

DGX agent

arXiv:2606.03203v1 Announce Type: new Abstract: Computer-use agents could automate repetitive screen-based clinical work, but their reliability in medical graphical user interfaces remains largely unv

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

MemoGen: Can Past Experience Improve Future Text-to-Image Generation?

DGX agent

arXiv:2606.03243v1 Announce Type: new Abstract: Modern text-to-image models have achieved strong visual synthesis, yet remain unreliable when prompts require implicit visual constraints, relational re

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Message Tuning Outshines Graph Prompt Tuning: A Prismatic Space Perspective

DGX agent

arXiv:2606.03290v1 Announce Type: cross Abstract: Graph Foundation Models (GFMs), built upon the Pre-training and Adaptation paradigm, have emerged as a research hotspot in graph learning. For GNN-bas

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

MiniMax M3 arrives with MiniMax Sparse Attention (MSA), 15.6x faster decoding at 1M tokens. We're partnering with @MiniMax_AI to power the i…

DGX agent

MiniMax M3 arrives with MiniMax Sparse Attention (MSA), 15.6x faster decoding at 1M tokens. We're partnering with @MiniMax_AI to power the inference behind this week's launch. Head to http://minimax.i

model-releasesfireworks-ai--x
3 Jun 2026
Model Releases

Mixed-Modality Dual Face-Hair Retrieval

DGX agent

arXiv:2606.03470v1 Announce Type: new Abstract: We introduce Dual Face-Hair Retrieval (DFHR), a new mixed-modality dual-reference task in image retrieval where a query consists of a face image specify

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

DGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Multi-Modal Machine Learning for Breast Cancer Recurrence Prediction

DGX agent

arXiv:2606.02892v1 Announce Type: new Abstract: Breast cancer recurrence, a leading cause of long-term mortality among survivors, requires timely and accurate risk assessment to guide follow-up care a

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Multi^2: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments

DGX agent

arXiv:2606.03698v1 Announce Type: new Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynam

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Multilingual Unlearning in LLMs: Transfer, Dynamics, and Reversibility

DGX agent

arXiv:2606.03291v1 Announce Type: new Abstract: Large language models (LLMs) can memorize sensitive facts, motivating unlearning methods that remove targeted knowledge without costly retraining. Howev

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

MultiTurnPSB: Evaluating Multi-Turn Jailbreak Attacks an dClassifier-Based Defenses for Medical AI Safety

DGX agent

arXiv:2606.02630v1 Announce Type: cross Abstract: Patient-facing medical chatbots are commonly evaluated on single-turn prompts, yet real users push back after refusals, add urgency, and invoke author

model-releasesarxiv-cs-ai
3 Jun 2026
← Previous
1…223224225226227…476
Next →