AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
11 Jun 2026

Parameter-Efficient Adapter Tuning for Tabular-Image Multimodal Learning

Model ReleasesDGX agent

arXiv:2606.11682v1 Announce Type: new Abstract: Tabular-image multimodal learning aims to improve predictive modeling by jointly using structured tabular attributes and visual data. Although pretraine

ParseFixer: An Agentic Framework for Document Parsing via Selective Multimodal Correction

Model ReleasesDGX agent

arXiv:2606.11977v1 Announce Type: new Abstract: In this report, we present our third-place solution for the DataMFM Challenge Track 1: Document Parsing. This track requires models to recover structure

Phi-Actor-Critic: Steering General-Sum Games to Pareto-Efficient Correlated Equilibria

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.11284v1 Announce Type: cross Abstract: Real-world multi-agent systems, from traffic coordination to resource allocation, are often modeled as general-sum games where individual incentives c

Robustness of Mixtures of Experts to Feature Noise

Model ReleasesDGX agent

arXiv:2601.14792v2 Announce Type: replace Abstract: Despite their practical success, it remains unclear why Mixture of Experts (MoE) models can outperform dense networks beyond sheer parameter scaling

SAFER-Nav: Enhancing Safety for Visual Robot Navigation via Segmentation-Aware Fine-Tuning

SafetyDGX agent

arXiv:2606.11636v1 Announce Type: new Abstract: Vision-based navigation models, particularly foundation models, generate viable trajectories from RGB observations alone. However, even state-of-the-art

SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment

SafetyDGX agent

arXiv:2606.11512v1 Announce Type: new Abstract: Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's samp

Steering the Noise: Turning Random Perturbations into Effective Descent for Memory-Efficient LLM Fine-Tuning

Model ReleasesDGX agent

arXiv:2601.04710v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) achieves strong performance but is often limited by the memory overhead of backpropagation. Zeroth-order (Z

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes

AgentsDGX agent

arXiv:2606.11470v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across natural language processing tasks, yet reliable reasoning remains an open challenge

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

Model ReleasesDGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

Model ReleasesDGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

Weighted Random Dot Product Graphs

ResearchDGX agent

arXiv:2505.03649v4 Announce Type: replace-cross Abstract: Modeling of intricate relational patterns has become a cornerstone of contemporary statistical research and related data science fields. Netwo

10 Jun 2026

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

SafetyDGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

5% > 100%: Flatness Preference is All You Need for Multimodal Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10488v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods provide a streamlined and efficient tool for adapting large models to domain-specific multimodal downstre

A Comprehensive Survey of Direct Preference Optimization: Datasets, Theories, Variants, and Applications

SafetyDGX agent

arXiv:2410.15595v4 Announce Type: replace Abstract: With the rapid advancement of large language models (LLMs), aligning policy models with human preferences has become increasingly critical. Direct P

A Constrained Natural-Language Interface for Variational Multi-Physics Finite Element Simulations in FEniCS

Model ReleasesDGX agent

arXiv:2606.10928v1 Announce Type: cross Abstract: Large language models can reduce the manual effort required to set up finite element simulations, but they introduce reliability risks when generated

A History-Aware Visually Grounded Critic for Computer Use Agents

Model ReleasesDGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

Model ReleasesDGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

[AINews] Anthropic Claude Fable 5 — Mythos but Safe, with Controversial Terms

Model ReleasesDGX agent

This article from Latent Space discusses Anthropic's Claude Fable 5 model, examining its capabilities in handling creative and mythological content while maintaining safety guardrails, along with cove

An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs

Model ReleasesDGX agent

arXiv:2603.14463v2 Announce Type: replace Abstract: Adapting Large Language Models (LLMs) to high-stakes vertical domains like insurance presents a significant challenge: scenarios demand strict adher

BadRobot: Jailbreaking Embodied LLM Agents in the Physical World

Model ReleasesDGX agent

arXiv:2407.20242v5 Announce Type: replace-cross Abstract: Embodied AI represents systems where AI is integrated into physical entities. Large Language Model (LLM), which exhibits powerful language und

Beyond Memorization: Distinguishing Between Pattern-Based and Epistemic Reasoning in LLMs Using Epistemic Puzzles

Model ReleasesDGX agent

arXiv:2603.21350v2 Announce Type: replace Abstract: Epistemic reasoning requires agents to infer the state of the world from partial observations and information about other agents' knowledge. Prior w

Blurry Window Attention

ResearchDGX agent

arXiv:2606.09862v1 Announce Type: cross Abstract: The Softmax Attention operation in Transformer language models has a quadratic complexity in the sequence length and a growing state size in the form

CoTAL: Human-in-the-Loop Prompt Engineering for Generalizable Formative Assessment Scoring and Feedback

Model ReleasesDGX agent

arXiv:2504.02323v4 Announce Type: replace Abstract: Large language models (LLMs) have created new opportunities to assist teachers and support student learning. While researchers have explored various

Data-Driven Dynamic Assortment in Online Platforms: Learning about Two Sides

Model ReleasesDGX agent

arXiv:2606.11118v1 Announce Type: new Abstract: We study a dynamic assortment problem on a two-sided service platform with incomplete information and heterogeneous customers in a discrete-time setting

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation

Model ReleasesDGX agent

arXiv:2606.10142v1 Announce Type: new Abstract: Recent advances in 3D generation have led to substantial improvements in realism, controllability, and efficiency, yet the evaluation of 3D assets remai

Do VLMs Reason Like Engineers? A Benchmark and a Stage-wise Evaluation

Model ReleasesDGX agent

arXiv:2606.10833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) demonstrate strong performance on general multimodal reasoning benchmarks, yet their ability to perform engineering reason

Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing

Model ReleasesDGX agent

arXiv:2606.10223v1 Announce Type: cross Abstract: Attributing a synthetic utterance to its originating system remains an open challenge: closed-set models fail to reject unseen synthesizers and produc

Enabling Progressive Whole-slide Image Analysis with Multi-scale Pyramidal Network

Model ReleasesDGX agent

arXiv:2602.01951v2 Announce Type: replace Abstract: Multiple-instance Learning (MIL) is commonly used for computational pathology (CPath), where multi-scale features are essential for capturing both f

Fisher-Guided Progressive Parameter Selection for Adaptive Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10196v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) aims to adapt pretrained models with a small trainable parameter subset, however, most existing methods choose

Flaws in the LLM Automation Narrative

Model ReleasesDGX agent

arXiv:2606.11166v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly described as performing at the level of human experts on knowledge economy tasks. These claims are prima

GhazalBench: Evaluating LLM Understanding and Canonical Surface-Form Access in Persian Ghazals

Model ReleasesDGX agent

arXiv:2603.09979v2 Announce Type: replace Abstract: Persian poetry plays an active role in Iranian cultural practice, where verses by canonical poets such as Hafez are frequently quoted, paraphrased,

GRID: Scaling Task-Agnostic Inference in Continual Prompt Tuning

Model ReleasesDGX agent

arXiv:2507.14725v4 Announce Type: replace-cross Abstract: Prompt-based continual learning (CL) offers a parameter-efficient way to adapt large language models (LLMs) across task sequences. However, ex

LakeQA: An Exploratory QA Benchmark over a Million-Scale Data Lake

Model ReleasesDGX agent

arXiv:2606.10460v1 Announce Type: cross Abstract: Recent large language models (LLMs) have shown rapid progress in reading-based question answering (QA), where evidence is explicitly provided or can b

Machine Learning Methods for Studying Latent Neural Activity Dynamics

ResearchDGX agent

arXiv:2606.10530v1 Announce Type: cross Abstract: Recent developments in brain recording are driving a demand for machine learning tools capable of decoding the latent structure of large populations o

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

Model ReleasesDGX agent

arXiv:2606.10742v1 Announce Type: cross Abstract: External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However,

Microsoft restricts Claude Fable for employees over data retention concerns

Model ReleasesDGX agent

Anthropic released Claude Fable, its first Mythos-class AI model, yesterday and it's already causing concerns inside Microsoft. Sources tell me that Microsoft is limiting the use of Claude Fable 5 for

Mixtures of Neural Operators Reduce Active Complexity in Operator Learning

Model ReleasesDGX agent

arXiv:2404.09101v3 Announce Type: replace-cross Abstract: Operator-learning systems are not governed solely by total parameter count; for one query, the relevant bottleneck can be the model that must

Nonlinear Estimator: Dual Bayesian Affine Estimators for Parameter Learning

Model ReleasesDGX agent

arXiv:2606.10111v1 Announce Type: new Abstract: This paper presents a nonlinear parameter estimator for Wiener-type state-space models obtained as a fixed-point architecture that couples two affine mi

NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI

Model ReleasesDGX agent

Today, Google DeepMind released DiffusionGemma — an experimental open model built for exceptionally fast text generation. NVIDIA has optimized DiffusionGemma to run even faster across NVIDIA GeForce R

One-Step Residual Shifting Diffusion for Image Super-Resolution via Distillation

HardwareDGX agent

arXiv:2503.13358v5 Announce Type: replace Abstract: Diffusion models for super-resolution (SR) produce high-quality visual results but require expensive computational costs. Despite the development of

Online Self-Training for Co-Adaptation in Hierarchical Diffusion Policies

Model ReleasesDGX agent

arXiv:2603.05291v2 Announce Type: replace Abstract: Hierarchical policies decompose language-conditioned long-horizon robotic manipulation into a high-level planner and a low-level controller. However

Privacy-Preserving Credit Risk Prediction with Alternative Data

ApplicationsDGX agent

arXiv:2606.10333v1 Announce Type: new Abstract: Credit risk prediction is a critical problem in the consumer credit industry. Traditionally, financial institutions construct credit risk prediction mod

Quality Is Not a Safety Proxy Under Quantization

Model ReleasesDGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

Model ReleasesDGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

Model ReleasesDGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

SPACR: Single-Pass Adaptive Training of Uncertainty-Aware Conformal Regressors

ResearchDGX agent

arXiv:2606.10734v1 Announce Type: new Abstract: Conformal Prediction (CP) provides robust uncertainty guarantees for predictive models, but is typically applied post hoc, which misaligns model trainin

Stop Early, Spend Less: Hidden-State Probes as a Practical Recipe for Streaming Moderation of LLM Outputs

SafetyDGX agent

arXiv:2606.10487v1 Announce Type: cross Abstract: Deploying large language models in user-facing systems requires efficient output safety filtering. Existing approaches typically rely on a separate mo

Structure from Reasoning, Numbers from Search: On-Premise Open LLMs as Structural Priors for Coupled MIMO Controller Tuning

Model ReleasesDGX agent

arXiv:2606.11015v1 Announce Type: new Abstract: Tuning controllers for strongly coupled multi-input multi-output (MIMO) industrial processes is hard: decentralized classical auto-tuning ignores loop i

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction

ApplicationsDGX agent

arXiv:2606.10279v1 Announce Type: new Abstract: Supervised fine-tuning with synthetic rationale data is widely assumed to improve language model performance on clinical prediction tasks by teaching mo

Temporal Sheaf Neural Networks with Dynamic Orthogonal Transport

Model ReleasesDGX agent

arXiv:2606.10071v1 Announce Type: cross Abstract: We introduce Temporal Sheaf Neural Networks (TSNN), a temporal link prediction framework that equips each node with a time-varying orthogonal frame an

The Role of Feedback Alignment in Self-Distillation

SafetyDGX agent

arXiv:2606.11173v1 Announce Type: new Abstract: Conditioning a language model on additional context, such as feedback on a previous attempt, typically improves its response. Self-distillation trains t

Trainable Smooth-Rotation Transforms with Learned Channel Scales for LLM Quantization

Model ReleasesDGX agent

arXiv:2606.09927v1 Announce Type: cross Abstract: Post-training quantization (PTQ) is one of the most practical ways to reduce the serving cost of Large Language Models (LLMs), but activation quantiza

Training LLMs to Enforce Multi-Level Instruction Hierarchies via Gravity-Weighted Direct Preference Optimization

Model ReleasesDGX agent

arXiv:2606.10860v1 Announce Type: cross Abstract: Production LLMs receive instructions from sources with very different levels of trust, yet attend to every token with uniform architectural privilege.

V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions

Model ReleasesDGX agent

arXiv:2512.11995v2 Announce Type: replace-cross Abstract: While many vision-language models (VLMs) are developed to answer well-defined, straightforward questions with highly specified targets, as in

Whisfusion: Parallel ASR Decoding with Masked Diffusion

ResearchDGX agent

arXiv:2508.07048v2 Announce Type: replace-cross Abstract: Autoregressive (AR) encoder-decoder models dominate high-quality multilingual ASR, but their left-to-right decoders make inference latency sca

XtrAIn: Training-Guided Occlusion for Feature Attribution

Model ReleasesDGX agent

arXiv:2606.10877v1 Announce Type: cross Abstract: Occlusion-based attribution methods provide an intuitive way to estimate feature importance by perturbing input features and measuring the resulting c

9 Jun 2026

A Framework for Evaluating and Benchmarking Concept Drift Detection Methods

Model ReleasesDGX agent

arXiv:2606.07789v1 Announce Type: new Abstract: Data stream mining is fundamentally challenged by concept drift, where distributional changes can degrade model performance. Despite the proliferation o

A Unifying Framework for Concept-Based Representational Similarity

Model ReleasesDGX agent

arXiv:2606.09653v1 Announce Type: new Abstract: Learned representations across models and modalities often exhibit striking structural similarities, suggesting shared underlying concept decompositions

AgentCompile: An LLM-Guided Compiler for Direct CUDA Inference

Model ReleasesDGX agent

arXiv:2606.07665v1 Announce Type: cross Abstract: Transformer inference increasingly depends on specialized compiler and runtime support, but real model graphs still require semantic decisions about w

AI Scientists Are Only as Good as Their Evidence: A Stratified Ablation of Proprietary Data and Reasoning Skills in Drug-Asset Valuation

Model ReleasesDGX agent

arXiv:2606.09556v1 Announce Type: new Abstract: AI Scientist agents are often evaluated as if capability were mainly a function of model quality, prompting, or reasoning scaffolds. We test a different

← Previous
1…396397398399400…1053
Next →