AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
20 Jun 2026

GLM-5.2 is great at design (Opus level IMO). I am also starting to see great results with long-running tasks, too. How is this possible? I t…

TutorialsDGX agent

GLM-5.2 is great at design (Opus level IMO). I am also starting to see great results with long-running tasks, too. How is this possible? I think there are a few clever hacks. But I just came across th

19 Jun 2026

https://huggingface.co/Lightricks/LTX-2.3-22b-IC-LoRA-Cross-Eyed Also shoutout to u/urabewe on Reddit created the content

Local AiDGX agent

LTX-2.3-22b-IC-LoRA-Cross-Eyed is a LoRA (Low-Rank Adaptation) model developed by Lightricks for the LTX-2.3-22b video generation model, created by Reddit user urabewe to address cross-eye artifacts i

11 Jun 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AnchorEdit: Maintaining Temporal Consistency in Multi-turn Image Editing via Causal Memory

Model ReleasesDGX agent

arXiv:2606.11751v1 Announce Type: cross Abstract: Multi-turn image editing is essential for iterative design, yet current models often struggle with identity drift and error accumulation over successi

APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection

Model ReleasesDGX agent

arXiv:2606.11459v1 Announce Type: cross Abstract: Large Language Models are highly sensitive to prompt formulation, necessitating automatic prompt optimization to unlock their full potential. While ev

ATLAS: Active Theory Learning for Automated Science

AgentsDGX agent

arXiv:2606.12386v1 Announce Type: cross Abstract: Advancing scientific understanding through mechanistic modeling requires posing the right experimental questions to yield maximally informative data.

Autoregressive Direct Preference Optimization

ResearchDGX agent

arXiv:2602.09533v2 Announce Type: replace Abstract: Direct preference optimization (DPO) has emerged as a promising approach for aligning large language models (LLMs) with human preferences. However,

CoVR-R:Reason-Aware Composed Video Retrieval

Model ReleasesDGX agent

arXiv:2603.20190v2 Announce Type: replace Abstract: Composed Video Retrieval (CoVR) aims to find a target video given a reference video and a textual modification. Prior work assumes the modification

CRUMB: Efficient Prior Fitted Network Inference via Distributionally Matched Context Batching

Model ReleasesDGX agent

arXiv:2606.11473v1 Announce Type: cross Abstract: Prior-fitted networks (PFNs) are a promising class of tabular foundation models that perform in-context learning, whereby the entire labelled training

Energy-Efficient On-Device RAG on a Mobile NPU: System Design and Benchmark on Snapdragon X Elite

Model ReleasesDGX agent

arXiv:2606.11257v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) pipelines are compute-intensive, combining embedding, retrieval, reranking, and large language model (LLM) generati

Exploration Structure in LLM Agents for Multi-File Change Localization

Model ReleasesDGX agent

arXiv:2606.11976v1 Announce Type: cross Abstract: Software engineering tools increasingly rely on LLM based agents to localize files to change to resolve a software issue. Most AI agents explore repos

Flow Matching with In-Context Priors for Out-of-Distribution Brain Dynamics

ResearchDGX agent

arXiv:2606.11833v1 Announce Type: new Abstract: Flow matching and diffusion models enable conditional generation across domains ranging from images to proteins, with recent extensions to out-of-distri

Grounding Computer Use Agents on Human Demonstrations

Model ReleasesDGX agent

arXiv:2511.07332v2 Announce Type: replace-cross Abstract: Building reliable computer-use agents requires grounding: accurately connecting natural language instructions to the correct on-screen element

Hey Chat, Can You Teach Me? Structuring Socratic Dialogue for Human Learning in the Wild

SafetyDGX agent

arXiv:2606.11744v1 Announce Type: cross Abstract: Large language models are now widely used for everyday learning, but the underlying interactions are typically unstructured chats rather than followin

How Auxiliary Reasoning Unleashes GUI Grounding in VLMs

Model ReleasesDGX agent

arXiv:2509.11548v2 Announce Type: replace Abstract: Graphical user interface (GUI) grounding is a fundamental task for building GUI agents. However, general vision-language models (VLMs) struggle with

Loss Landscape Diagnosis for Gradient-Based Gray-Scott System Inversion: Disentangling the Roles of PINN Components

Model ReleasesDGX agent

arXiv:2606.11258v1 Announce Type: new Abstract: Gradient-based inversion of reaction-diffusion systems is typically approached via surrogate models or physics-informed neural networks (PINNs), while t

MPC-Patch-Bench: Security-Aware LLM Code Patch for Multi-Party Computation

Model ReleasesDGX agent

arXiv:2606.11416v1 Announce Type: cross Abstract: Repository-level benchmarks for evaluating Large Language Model (LLM) code repair on Secure Multi-Party Computation (MPC) software do not yet exist, a

OSCS-SupCon: Orthogonal Sigmoid-based Common and Style Supervised Contrastive Learning for Robust Feature Disentanglement

Model ReleasesDGX agent

arXiv:2606.11233v1 Announce Type: new Abstract: Supervised Contrastive Learning (SupCon) has achieved strong performance by explicitly modeling pairwise relationships among samples. However, existing

Parameter-Efficient Adapter Tuning for Tabular-Image Multimodal Learning

Model ReleasesDGX agent

arXiv:2606.11682v1 Announce Type: new Abstract: Tabular-image multimodal learning aims to improve predictive modeling by jointly using structured tabular attributes and visual data. Although pretraine

ParseFixer: An Agentic Framework for Document Parsing via Selective Multimodal Correction

Model ReleasesDGX agent

arXiv:2606.11977v1 Announce Type: new Abstract: In this report, we present our third-place solution for the DataMFM Challenge Track 1: Document Parsing. This track requires models to recover structure

Phi-Actor-Critic: Steering General-Sum Games to Pareto-Efficient Correlated Equilibria

Model ReleasesDGX agent

arXiv:2606.11284v1 Announce Type: cross Abstract: Real-world multi-agent systems, from traffic coordination to resource allocation, are often modeled as general-sum games where individual incentives c

Robustness of Mixtures of Experts to Feature Noise

Model ReleasesDGX agent

arXiv:2601.14792v2 Announce Type: replace Abstract: Despite their practical success, it remains unclear why Mixture of Experts (MoE) models can outperform dense networks beyond sheer parameter scaling

SAFER-Nav: Enhancing Safety for Visual Robot Navigation via Segmentation-Aware Fine-Tuning

SafetyDGX agent

arXiv:2606.11636v1 Announce Type: new Abstract: Vision-based navigation models, particularly foundation models, generate viable trajectories from RGB observations alone. However, even state-of-the-art

SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment

SafetyDGX agent

arXiv:2606.11512v1 Announce Type: new Abstract: Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's samp

Steering the Noise: Turning Random Perturbations into Effective Descent for Memory-Efficient LLM Fine-Tuning

Model ReleasesDGX agent

arXiv:2601.04710v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) achieves strong performance but is often limited by the memory overhead of backpropagation. Zeroth-order (Z

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes

AgentsDGX agent

arXiv:2606.11470v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across natural language processing tasks, yet reliable reasoning remains an open challenge

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

Model ReleasesDGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

Model ReleasesDGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

Weighted Random Dot Product Graphs

ResearchDGX agent

arXiv:2505.03649v4 Announce Type: replace-cross Abstract: Modeling of intricate relational patterns has become a cornerstone of contemporary statistical research and related data science fields. Netwo

10 Jun 2026

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

SafetyDGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

5% > 100%: Flatness Preference is All You Need for Multimodal Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10488v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods provide a streamlined and efficient tool for adapting large models to domain-specific multimodal downstre

A Comprehensive Survey of Direct Preference Optimization: Datasets, Theories, Variants, and Applications

SafetyDGX agent

arXiv:2410.15595v4 Announce Type: replace Abstract: With the rapid advancement of large language models (LLMs), aligning policy models with human preferences has become increasingly critical. Direct P

A Constrained Natural-Language Interface for Variational Multi-Physics Finite Element Simulations in FEniCS

Model ReleasesDGX agent

arXiv:2606.10928v1 Announce Type: cross Abstract: Large language models can reduce the manual effort required to set up finite element simulations, but they introduce reliability risks when generated

A History-Aware Visually Grounded Critic for Computer Use Agents

Model ReleasesDGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

Model ReleasesDGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

[AINews] Anthropic Claude Fable 5 — Mythos but Safe, with Controversial Terms

Model ReleasesDGX agent

This article from Latent Space discusses Anthropic's Claude Fable 5 model, examining its capabilities in handling creative and mythological content while maintaining safety guardrails, along with cove

An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs

Model ReleasesDGX agent

arXiv:2603.14463v2 Announce Type: replace Abstract: Adapting Large Language Models (LLMs) to high-stakes vertical domains like insurance presents a significant challenge: scenarios demand strict adher

BadRobot: Jailbreaking Embodied LLM Agents in the Physical World

Model ReleasesDGX agent

arXiv:2407.20242v5 Announce Type: replace-cross Abstract: Embodied AI represents systems where AI is integrated into physical entities. Large Language Model (LLM), which exhibits powerful language und

Beyond Memorization: Distinguishing Between Pattern-Based and Epistemic Reasoning in LLMs Using Epistemic Puzzles

Model ReleasesDGX agent

arXiv:2603.21350v2 Announce Type: replace Abstract: Epistemic reasoning requires agents to infer the state of the world from partial observations and information about other agents' knowledge. Prior w

Blurry Window Attention

ResearchDGX agent

arXiv:2606.09862v1 Announce Type: cross Abstract: The Softmax Attention operation in Transformer language models has a quadratic complexity in the sequence length and a growing state size in the form

CoTAL: Human-in-the-Loop Prompt Engineering for Generalizable Formative Assessment Scoring and Feedback

Model ReleasesDGX agent

arXiv:2504.02323v4 Announce Type: replace Abstract: Large language models (LLMs) have created new opportunities to assist teachers and support student learning. While researchers have explored various

Data-Driven Dynamic Assortment in Online Platforms: Learning about Two Sides

Model ReleasesDGX agent

arXiv:2606.11118v1 Announce Type: new Abstract: We study a dynamic assortment problem on a two-sided service platform with incomplete information and heterogeneous customers in a discrete-time setting

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation

Model ReleasesDGX agent

arXiv:2606.10142v1 Announce Type: new Abstract: Recent advances in 3D generation have led to substantial improvements in realism, controllability, and efficiency, yet the evaluation of 3D assets remai

Do VLMs Reason Like Engineers? A Benchmark and a Stage-wise Evaluation

Model ReleasesDGX agent

arXiv:2606.10833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) demonstrate strong performance on general multimodal reasoning benchmarks, yet their ability to perform engineering reason

Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing

Model ReleasesDGX agent

arXiv:2606.10223v1 Announce Type: cross Abstract: Attributing a synthetic utterance to its originating system remains an open challenge: closed-set models fail to reject unseen synthesizers and produc

Enabling Progressive Whole-slide Image Analysis with Multi-scale Pyramidal Network

Model ReleasesDGX agent

arXiv:2602.01951v2 Announce Type: replace Abstract: Multiple-instance Learning (MIL) is commonly used for computational pathology (CPath), where multi-scale features are essential for capturing both f

Fisher-Guided Progressive Parameter Selection for Adaptive Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10196v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) aims to adapt pretrained models with a small trainable parameter subset, however, most existing methods choose

Flaws in the LLM Automation Narrative

Model ReleasesDGX agent

arXiv:2606.11166v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly described as performing at the level of human experts on knowledge economy tasks. These claims are prima

GhazalBench: Evaluating LLM Understanding and Canonical Surface-Form Access in Persian Ghazals

Model ReleasesDGX agent

arXiv:2603.09979v2 Announce Type: replace Abstract: Persian poetry plays an active role in Iranian cultural practice, where verses by canonical poets such as Hafez are frequently quoted, paraphrased,

GRID: Scaling Task-Agnostic Inference in Continual Prompt Tuning

Model ReleasesDGX agent

arXiv:2507.14725v4 Announce Type: replace-cross Abstract: Prompt-based continual learning (CL) offers a parameter-efficient way to adapt large language models (LLMs) across task sequences. However, ex

LakeQA: An Exploratory QA Benchmark over a Million-Scale Data Lake

Model ReleasesDGX agent

arXiv:2606.10460v1 Announce Type: cross Abstract: Recent large language models (LLMs) have shown rapid progress in reading-based question answering (QA), where evidence is explicitly provided or can b

Machine Learning Methods for Studying Latent Neural Activity Dynamics

ResearchDGX agent

arXiv:2606.10530v1 Announce Type: cross Abstract: Recent developments in brain recording are driving a demand for machine learning tools capable of decoding the latent structure of large populations o

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

Model ReleasesDGX agent

arXiv:2606.10742v1 Announce Type: cross Abstract: External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However,

Microsoft restricts Claude Fable for employees over data retention concerns

Model ReleasesDGX agent

Anthropic released Claude Fable, its first Mythos-class AI model, yesterday and it's already causing concerns inside Microsoft. Sources tell me that Microsoft is limiting the use of Claude Fable 5 for

Mixtures of Neural Operators Reduce Active Complexity in Operator Learning

Model ReleasesDGX agent

arXiv:2404.09101v3 Announce Type: replace-cross Abstract: Operator-learning systems are not governed solely by total parameter count; for one query, the relevant bottleneck can be the model that must

Nonlinear Estimator: Dual Bayesian Affine Estimators for Parameter Learning

Model ReleasesDGX agent

arXiv:2606.10111v1 Announce Type: new Abstract: This paper presents a nonlinear parameter estimator for Wiener-type state-space models obtained as a fixed-point architecture that couples two affine mi

NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI

Model ReleasesDGX agent

Today, Google DeepMind released DiffusionGemma — an experimental open model built for exceptionally fast text generation. NVIDIA has optimized DiffusionGemma to run even faster across NVIDIA GeForce R

One-Step Residual Shifting Diffusion for Image Super-Resolution via Distillation

HardwareDGX agent

arXiv:2503.13358v5 Announce Type: replace Abstract: Diffusion models for super-resolution (SR) produce high-quality visual results but require expensive computational costs. Despite the development of

Online Self-Training for Co-Adaptation in Hierarchical Diffusion Policies

Model ReleasesDGX agent

arXiv:2603.05291v2 Announce Type: replace Abstract: Hierarchical policies decompose language-conditioned long-horizon robotic manipulation into a high-level planner and a low-level controller. However

Privacy-Preserving Credit Risk Prediction with Alternative Data

ApplicationsDGX agent

arXiv:2606.10333v1 Announce Type: new Abstract: Credit risk prediction is a critical problem in the consumer credit industry. Traditionally, financial institutions construct credit risk prediction mod

Quality Is Not a Safety Proxy Under Quantization

Model ReleasesDGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

← Previous
1…395396397398399…1051
Next →