AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Model Releases

Parameter-Efficient Adapter Tuning for Tabular-Image Multimodal Learning

DGX agent

arXiv:2606.11682v1 Announce Type: new Abstract: Tabular-image multimodal learning aims to improve predictive modeling by jointly using structured tabular attributes and visual data. Although pretraine

model-releasesarxiv-cs-cv
11 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ParseFixer: An Agentic Framework for Document Parsing via Selective Multimodal Correction

DGX agent

arXiv:2606.11977v1 Announce Type: new Abstract: In this report, we present our third-place solution for the DataMFM Challenge Track 1: Document Parsing. This track requires models to recover structure

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Phi-Actor-Critic: Steering General-Sum Games to Pareto-Efficient Correlated Equilibria

DGX agent

arXiv:2606.11284v1 Announce Type: cross Abstract: Real-world multi-agent systems, from traffic coordination to resource allocation, are often modeled as general-sum games where individual incentives c

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Robustness of Mixtures of Experts to Feature Noise

DGX agent

arXiv:2601.14792v2 Announce Type: replace Abstract: Despite their practical success, it remains unclear why Mixture of Experts (MoE) models can outperform dense networks beyond sheer parameter scaling

model-releasesarxiv-cs-lg
11 Jun 2026
Safety

SAFER-Nav: Enhancing Safety for Visual Robot Navigation via Segmentation-Aware Fine-Tuning

DGX agent

arXiv:2606.11636v1 Announce Type: new Abstract: Vision-based navigation models, particularly foundation models, generate viable trajectories from RGB observations alone. However, even state-of-the-art

safetyarxiv-cs-ro
11 Jun 2026
Safety

SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment

DGX agent

arXiv:2606.11512v1 Announce Type: new Abstract: Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's samp

safetyarxiv-cs-cl
11 Jun 2026
Model Releases

Steering the Noise: Turning Random Perturbations into Effective Descent for Memory-Efficient LLM Fine-Tuning

DGX agent

arXiv:2601.04710v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) achieves strong performance but is often limited by the memory overhead of backpropagation. Zeroth-order (Z

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes

DGX agent

arXiv:2606.11470v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across natural language processing tasks, yet reliable reasoning remains an open challenge

agentsarxiv-cs-cl
11 Jun 2026
Model Releases

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

DGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

DGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

model-releasesarxiv-cs-cv
11 Jun 2026
Research

Weighted Random Dot Product Graphs

DGX agent

arXiv:2505.03649v4 Announce Type: replace-cross Abstract: Modeling of intricate relational patterns has become a cornerstone of contemporary statistical research and related data science fields. Netwo

researcharxiv-cs-lg
11 Jun 2026
Safety

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

DGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

5% > 100%: Flatness Preference is All You Need for Multimodal Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2606.10488v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods provide a streamlined and efficient tool for adapting large models to domain-specific multimodal downstre

model-releasesarxiv-cs-cv
10 Jun 2026
Safety

A Comprehensive Survey of Direct Preference Optimization: Datasets, Theories, Variants, and Applications

DGX agent

arXiv:2410.15595v4 Announce Type: replace Abstract: With the rapid advancement of large language models (LLMs), aligning policy models with human preferences has become increasingly critical. Direct P

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

A Constrained Natural-Language Interface for Variational Multi-Physics Finite Element Simulations in FEniCS

DGX agent

arXiv:2606.10928v1 Announce Type: cross Abstract: Large language models can reduce the manual effort required to set up finite element simulations, but they introduce reliability risks when generated

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

A History-Aware Visually Grounded Critic for Computer Use Agents

DGX agent

arXiv:2606.11078v1 Announce Type: new Abstract: Various test-time interventions for Computer Use Agents (CUAs), including critic models, have been developed to improve performance through pre-executio

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

DGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

[AINews] Anthropic Claude Fable 5 — Mythos but Safe, with Controversial Terms

DGX agent

This article from Latent Space discusses Anthropic's Claude Fable 5 model, examining its capabilities in handling creative and mythological content while maintaining safety guardrails, along with cove

model-releaseslatent-space
10 Jun 2026
Model Releases

An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs

DGX agent

arXiv:2603.14463v2 Announce Type: replace Abstract: Adapting Large Language Models (LLMs) to high-stakes vertical domains like insurance presents a significant challenge: scenarios demand strict adher

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

BadRobot: Jailbreaking Embodied LLM Agents in the Physical World

DGX agent

arXiv:2407.20242v5 Announce Type: replace-cross Abstract: Embodied AI represents systems where AI is integrated into physical entities. Large Language Model (LLM), which exhibits powerful language und

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Beyond Memorization: Distinguishing Between Pattern-Based and Epistemic Reasoning in LLMs Using Epistemic Puzzles

DGX agent

arXiv:2603.21350v2 Announce Type: replace Abstract: Epistemic reasoning requires agents to infer the state of the world from partial observations and information about other agents' knowledge. Prior w

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Blurry Window Attention

DGX agent

arXiv:2606.09862v1 Announce Type: cross Abstract: The Softmax Attention operation in Transformer language models has a quadratic complexity in the sequence length and a growing state size in the form

researcharxiv-cs-ai
10 Jun 2026
Model Releases

CoTAL: Human-in-the-Loop Prompt Engineering for Generalizable Formative Assessment Scoring and Feedback

DGX agent

arXiv:2504.02323v4 Announce Type: replace Abstract: Large language models (LLMs) have created new opportunities to assist teachers and support student learning. While researchers have explored various

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Data-Driven Dynamic Assortment in Online Platforms: Learning about Two Sides

DGX agent

arXiv:2606.11118v1 Announce Type: new Abstract: We study a dynamic assortment problem on a two-sided service platform with incomplete information and heterogeneous customers in a discrete-time setting

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation

DGX agent

arXiv:2606.10142v1 Announce Type: new Abstract: Recent advances in 3D generation have led to substantial improvements in realism, controllability, and efficiency, yet the evaluation of 3D assets remai

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Do VLMs Reason Like Engineers? A Benchmark and a Stage-wise Evaluation

DGX agent

arXiv:2606.10833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) demonstrate strong performance on general multimodal reasoning benchmarks, yet their ability to perform engineering reason

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Dual-Branch Gated Fusion for Open-Set Audio Deepfake Source Tracing

DGX agent

arXiv:2606.10223v1 Announce Type: cross Abstract: Attributing a synthetic utterance to its originating system remains an open challenge: closed-set models fail to reject unseen synthesizers and produc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Enabling Progressive Whole-slide Image Analysis with Multi-scale Pyramidal Network

DGX agent

arXiv:2602.01951v2 Announce Type: replace Abstract: Multiple-instance Learning (MIL) is commonly used for computational pathology (CPath), where multi-scale features are essential for capturing both f

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Fisher-Guided Progressive Parameter Selection for Adaptive Fine-Tuning

DGX agent

arXiv:2606.10196v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) aims to adapt pretrained models with a small trainable parameter subset, however, most existing methods choose

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Flaws in the LLM Automation Narrative

DGX agent

arXiv:2606.11166v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly described as performing at the level of human experts on knowledge economy tasks. These claims are prima

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

GhazalBench: Evaluating LLM Understanding and Canonical Surface-Form Access in Persian Ghazals

DGX agent

arXiv:2603.09979v2 Announce Type: replace Abstract: Persian poetry plays an active role in Iranian cultural practice, where verses by canonical poets such as Hafez are frequently quoted, paraphrased,

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

GRID: Scaling Task-Agnostic Inference in Continual Prompt Tuning

DGX agent

arXiv:2507.14725v4 Announce Type: replace-cross Abstract: Prompt-based continual learning (CL) offers a parameter-efficient way to adapt large language models (LLMs) across task sequences. However, ex

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

LakeQA: An Exploratory QA Benchmark over a Million-Scale Data Lake

DGX agent

arXiv:2606.10460v1 Announce Type: cross Abstract: Recent large language models (LLMs) have shown rapid progress in reading-based question answering (QA), where evidence is explicitly provided or can b

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Machine Learning Methods for Studying Latent Neural Activity Dynamics

DGX agent

arXiv:2606.10530v1 Announce Type: cross Abstract: Recent developments in brain recording are driving a demand for machine learning tools capable of decoding the latent structure of large populations o

researcharxiv-cs-ai
10 Jun 2026
Model Releases

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

DGX agent

arXiv:2606.10742v1 Announce Type: cross Abstract: External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However,

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Microsoft restricts Claude Fable for employees over data retention concerns

DGX agent

Anthropic released Claude Fable, its first Mythos-class AI model, yesterday and it's already causing concerns inside Microsoft. Sources tell me that Microsoft is limiting the use of Claude Fable 5 for

model-releasesthe-verge-ai
10 Jun 2026
Model Releases

Mixtures of Neural Operators Reduce Active Complexity in Operator Learning

DGX agent

arXiv:2404.09101v3 Announce Type: replace-cross Abstract: Operator-learning systems are not governed solely by total parameter count; for one query, the relevant bottleneck can be the model that must

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Nonlinear Estimator: Dual Bayesian Affine Estimators for Parameter Learning

DGX agent

arXiv:2606.10111v1 Announce Type: new Abstract: This paper presents a nonlinear parameter estimator for Wiener-type state-space models obtained as a fixed-point architecture that couples two affine mi

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI

DGX agent

Today, Google DeepMind released DiffusionGemma — an experimental open model built for exceptionally fast text generation. NVIDIA has optimized DiffusionGemma to run even faster across NVIDIA GeForce R

model-releasesnvidia-blog
10 Jun 2026
Hardware

One-Step Residual Shifting Diffusion for Image Super-Resolution via Distillation

DGX agent

arXiv:2503.13358v5 Announce Type: replace Abstract: Diffusion models for super-resolution (SR) produce high-quality visual results but require expensive computational costs. Despite the development of

hardwarearxiv-cs-cv
10 Jun 2026
Model Releases

Online Self-Training for Co-Adaptation in Hierarchical Diffusion Policies

DGX agent

arXiv:2603.05291v2 Announce Type: replace Abstract: Hierarchical policies decompose language-conditioned long-horizon robotic manipulation into a high-level planner and a low-level controller. However

model-releasesarxiv-cs-ro
10 Jun 2026
Applications

Privacy-Preserving Credit Risk Prediction with Alternative Data

DGX agent

arXiv:2606.10333v1 Announce Type: new Abstract: Credit risk prediction is a critical problem in the consumer credit industry. Traditionally, financial institutions construct credit risk prediction mod

applicationsarxiv-cs-lg
10 Jun 2026
Model Releases

Quality Is Not a Safety Proxy Under Quantization

DGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

DGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

DGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

model-releasesarxiv-cs-ro
10 Jun 2026
Research

SPACR: Single-Pass Adaptive Training of Uncertainty-Aware Conformal Regressors

DGX agent

arXiv:2606.10734v1 Announce Type: new Abstract: Conformal Prediction (CP) provides robust uncertainty guarantees for predictive models, but is typically applied post hoc, which misaligns model trainin

researcharxiv-cs-lg
10 Jun 2026
Safety

Stop Early, Spend Less: Hidden-State Probes as a Practical Recipe for Streaming Moderation of LLM Outputs

DGX agent

arXiv:2606.10487v1 Announce Type: cross Abstract: Deploying large language models in user-facing systems requires efficient output safety filtering. Existing approaches typically rely on a separate mo

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Structure from Reasoning, Numbers from Search: On-Premise Open LLMs as Structural Priors for Coupled MIMO Controller Tuning

DGX agent

arXiv:2606.11015v1 Announce Type: new Abstract: Tuning controllers for strongly coupled multi-input multi-output (MIMO) industrial processes is hard: decentralized classical auto-tuning ignores loop i

model-releasesarxiv-cs-ai
10 Jun 2026
← Previous
1…518519520521522…1369
Next →