AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Model Releases

Deep Learning-Based Sign Language Recognition from Videos and Cross-Lingual Translation to Indian Vernaculars

DGX agent

arXiv:2606.22494v1 Announce Type: cross Abstract: Sign language is a primary mode of communication for the global deaf and hard-of-hearing community, yet automated tools that recognize sign gestures f

model-releasesarxiv-cs-lg
23 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Deep Learning for Individual Heterogeneity

DGX agent

arXiv:2010.14694v4 Announce Type: replace-cross Abstract: This paper integrates deep neural networks (DNNs) into structural models to increase flexibility and capture rich heterogeneity while preservi

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Demystifying Numerical Instability in LLM Inference: Achieving Reproducible Inference for Mission-Critical Tasks with HEAL

DGX agent

arXiv:2606.21023v1 Announce Type: new Abstract: As Large Language Models (LLMs) deploy into mission-critical domains (e.g., finance, medicine, and law), output reproducibility has become a strict syst

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

DrugBench: Evaluating AI Control Protocols for Medication Harm Mitigation

DGX agent

arXiv:2606.20663v1 Announce Type: cross Abstract: Large Language Models have the potential to expand and improve the access to clinical information by enabling new ways of interacting with medical kno

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Dynamics, stability, and energy efficiency of an energy-recycling rimless wheel with spring-clutch legs

DGX agent

arXiv:2606.22073v1 Announce Type: new Abstract: This paper proposes an energy-recycling rimless wheel with spring-clutch legs. The proposed mechanism uses a lockable clutch to store part of the impact

model-releasesarxiv-cs-ro
23 Jun 2026
Safety

Efficient Reinforcement Finetuning via Adaptive Curriculum Learning

DGX agent

arXiv:2504.05520v4 Announce Type: replace Abstract: Reinforcement finetuning (RFT) has shown great potential for enhancing the mathematical reasoning capabilities of large language models (LLMs), but

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Flowing With Purpose: Latent Action Guided Flow Matching Policies For Robotic Manipulation

DGX agent

arXiv:2606.23420v1 Announce Type: new Abstract: Flow matching has recently become a new standard for behavior cloning in robotic manipulation. However, state-of-the-art flow matching policies suffer f

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

GeoFidelity-Bench: Evaluating Segment-Level Geographic Fidelity in Text-to-Image Street-View Generation

DGX agent

arXiv:2606.23669v1 Announce Type: new Abstract: Text-to-image models can generate visually plausible city streets, but whether their outputs correspond to a requested road segment rather than a generi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

GRAG: Generic Response-Augmented Generation Framework for Personalized Conversational Systems

DGX agent

arXiv:2606.21097v1 Announce Type: cross Abstract: Deploying highly capable personalized conversational agents in resource-constrained or privacy-sensitive environments remains a significant challenge.

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

In-Context Molecular Property Prediction with LLMs: A Blinding Study on Memorization and Knowledge Conflicts

DGX agent

arXiv:2603.25857v2 Announce Type: replace Abstract: The capabilities of large language models (LLMs) have expanded beyond natural language processing to scientific prediction tasks, including molecula

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Introducing Mistral OCR 4. It creates structure with bounding boxes, block classification, and inline confidence scores in 170 languages. 🧵…

DGX agent

Mistral OCR 4 is an optical character recognition model that extracts text with structural information including bounding boxes and block classification across 170 languages. The system provides inlin

model-releasesarthur-mensch--x
23 Jun 2026
Model Releases

Is Our Benchmark Enough? An Analysis of Continual Learning for MLLMs

DGX agent

arXiv:2606.20961v1 Announce Type: new Abstract: Continual adaptation is essential for multimodal large language models (MLLMs) deployed across evolving domains, but the state-of-the-art MR-LoRA method

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Learning Bug Context for PyTorch-to-JAX Translation with LLMs

DGX agent

arXiv:2510.09898v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong performance on code translation between widely used programming languages. However, translation becom

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

MoECodec: Image Compression for joint human and machine perception via Mixture-of-Experts

DGX agent

arXiv:2606.21033v1 Announce Type: cross Abstract: Image compression for machines calls for a unified codec that serves multiple downstream vision tasks. Existing approaches either adopt task-specific

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

MotionHalluc: Diagnosing Kinematic Hallucinations in Fine-Grained Motion Reasoning

DGX agent

arXiv:2606.23061v1 Announce Type: new Abstract: Motion instruction generation in cross-video comparison aims to produce corrective feedback that describes the differences between a query and a referen

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Open Problem: Is AdamW Effective Under Heavy-Tailed Noise?

DGX agent

arXiv:2606.23676v1 Announce Type: new Abstract: AdamW is the de facto optimizer for training large language models (LLMs), yet the theory behind it still lives mostly in finite-variance regimes. This

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Physics-Guided Dual-Stream Heterogeneous Graph Neural Network for Predicting Full-Field Structural Response of Stiffened Panels

DGX agent

arXiv:2606.20916v1 Announce Type: new Abstract: Iterative design and optimization of large, complex structures require fast and accurate prediction of stress, displacement, and other fields. Finite el

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Priority-Aware Learning-Unlearning Correction for Dynamic Decentralized LoRA Fine-Tuning

DGX agent

arXiv:2606.22878v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed at the network edge to provide pervasive generative AI services, decentralized federated learn

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

READ More than What You See: Reinforcement Learning for Accurate and Coherent Audio Description Generations

DGX agent

arXiv:2606.22766v1 Announce Type: new Abstract: Audio Description aims to generate concise narrations of essential visual content in audio-visual media for blind and low-vision audiences. Existing met

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Revealing the Pitfalls and Re-Evaluating the Advancement of Heterophilic Graph Learning

DGX agent

arXiv:2409.05755v3 Announce Type: replace Abstract: Over the past decade, Graph Neural Networks (GNNs) have achieved great success on machine learning tasks with relational data. However, recent studi

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RoverDevKit: An open, physics-grounded tradespace toolkit for conceptual design of lunar micro-rovers

DGX agent

arXiv:2606.21755v1 Announce Type: new Abstract: Pre-Phase-A design of lunar micro-rovers is dominated by tightly coupled mobility, power, thermal, and mass trades, yet conceptual-design tooling for th

model-releasesarxiv-cs-ro
23 Jun 2026
Research

S^2VG: 3D Stereoscopic and Spatial Video Generation via Denoising Frame Matrix

DGX agent

arXiv:2508.08048v2 Announce Type: replace Abstract: While video generation models excel at producing high-quality monocular videos, generating 3D stereoscopic and spatial videos for immersive applicat

researcharxiv-cs-cv
23 Jun 2026
Model Releases

Self-Improvement Can Self-Regress: The Rise-and-Collapse Failure Mode of LLM Self-Training

DGX agent

arXiv:2606.21090v1 Announce Type: cross Abstract: Self-improvement can self-regress. In REINFORCE post-training for code, a model can quickly improve on its optimized metric and then collapse within t

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Skill Coverage: A Test Adequacy Metric for Agent Skills

DGX agent

arXiv:2606.20659v1 Announce Type: cross Abstract: Agent skills encode reusable procedural knowledge that guides large language model agents across tasks and execution contexts. Existing evaluations pr

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

Subspace-Constrained Federated Learning with Low-Rank Adaptation

DGX agent

arXiv:2606.22724v1 Announce Type: new Abstract: Federated low-rank adaptation methods are attractive for fine-tuning large models under communication and privacy constraints, but heterogeneous client

local-aiarxiv-cs-lg
23 Jun 2026
Model Releases

Temporally Aware Densification for Dynamic 3D Gaussian Splatting

DGX agent

arXiv:2606.23212v1 Announce Type: new Abstract: Despite modeling temporal motion, dynamic 3D Gaussian Splatting (3DGS) methods still inherit a static densification strategy that is ill-suited for dyna

model-releasesarxiv-cs-cv
23 Jun 2026
Research

Tensor Train Decomposition-based 3D Implicit Full Waveform Inversion with Multi-scale Structural Similarity

DGX agent

arXiv:2606.22867v1 Announce Type: cross Abstract: Three-dimensional full waveform inversion (3DFWI) is a powerful technique for reconstructing high-resolution subsurface velocity models. However, its

researcharxiv-cs-lg
23 Jun 2026
Safety

The Pitfall of Scaling Up: Uncovering and Mitigating Popularity Bias Amplification in Scaling Transformer-based Recommenders

DGX agent

arXiv:2606.21911v1 Announce Type: cross Abstract: We identify a critical pitfall in scaling transformer-based sequential recommenders: while increasing model size improves recommendation accuracy, it

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents

DGX agent

arXiv:2606.22864v1 Announce Type: new Abstract: Hidden-state probing -- a linear classifier on a frozen vision-language model's internal activations -- has emerged as an attractive evaluation tool for

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

With agentic coding, complexity compounds in a mechanical way: unnecessary code ends up in the codebase, moves to the context window, degrad…

DGX agent

With agentic coding, complexity compounds in a mechanical way: unnecessary code ends up in the codebase, moves to the context window, degrades the model's reasoning abilities, leads to more unnecessar

model-releasesfrancois-chollet--x
23 Jun 2026
Tools

PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters

DGX agent

PP-OCRv6 is a multilingual optical character recognition (OCR) model family released by PaddlePaddle that supports 50 languages across multiple parameter sizes ranging from 1.5M to 34.5M. The model se

toolshugging-face
22 Jun 2026
Tutorials

GLM-5.2 is great at design (Opus level IMO). I am also starting to see great results with long-running tasks, too. How is this possible? I t…

DGX agent

GLM-5.2 is great at design (Opus level IMO). I am also starting to see great results with long-running tasks, too. How is this possible? I think there are a few clever hacks. But I just came across th

tutorialsdair-ai--x
20 Jun 2026
Local Ai

https://huggingface.co/Lightricks/LTX-2.3-22b-IC-LoRA-Cross-Eyed Also shoutout to u/urabewe on Reddit created the content

DGX agent

LTX-2.3-22b-IC-LoRA-Cross-Eyed is a LoRA (Low-Rank Adaptation) model developed by Lightricks for the LTX-2.3-22b video generation model, created by Reddit user urabewe to address cross-eye artifacts i

local-aicomfyui--x
19 Jun 2026
Model Releases

AnchorEdit: Maintaining Temporal Consistency in Multi-turn Image Editing via Causal Memory

DGX agent

arXiv:2606.11751v1 Announce Type: cross Abstract: Multi-turn image editing is essential for iterative design, yet current models often struggle with identity drift and error accumulation over successi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection

DGX agent

arXiv:2606.11459v1 Announce Type: cross Abstract: Large Language Models are highly sensitive to prompt formulation, necessitating automatic prompt optimization to unlock their full potential. While ev

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

ATLAS: Active Theory Learning for Automated Science

DGX agent

arXiv:2606.12386v1 Announce Type: cross Abstract: Advancing scientific understanding through mechanistic modeling requires posing the right experimental questions to yield maximally informative data.

agentsarxiv-cs-ai
11 Jun 2026
Research

Autoregressive Direct Preference Optimization

DGX agent

arXiv:2602.09533v2 Announce Type: replace Abstract: Direct preference optimization (DPO) has emerged as a promising approach for aligning large language models (LLMs) with human preferences. However,

researcharxiv-cs-ai
11 Jun 2026
Model Releases

CoVR-R:Reason-Aware Composed Video Retrieval

DGX agent

arXiv:2603.20190v2 Announce Type: replace Abstract: Composed Video Retrieval (CoVR) aims to find a target video given a reference video and a textual modification. Prior work assumes the modification

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

CRUMB: Efficient Prior Fitted Network Inference via Distributionally Matched Context Batching

DGX agent

arXiv:2606.11473v1 Announce Type: cross Abstract: Prior-fitted networks (PFNs) are a promising class of tabular foundation models that perform in-context learning, whereby the entire labelled training

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Energy-Efficient On-Device RAG on a Mobile NPU: System Design and Benchmark on Snapdragon X Elite

DGX agent

arXiv:2606.11257v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) pipelines are compute-intensive, combining embedding, retrieval, reranking, and large language model (LLM) generati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Exploration Structure in LLM Agents for Multi-File Change Localization

DGX agent

arXiv:2606.11976v1 Announce Type: cross Abstract: Software engineering tools increasingly rely on LLM based agents to localize files to change to resolve a software issue. Most AI agents explore repos

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Flow Matching with In-Context Priors for Out-of-Distribution Brain Dynamics

DGX agent

arXiv:2606.11833v1 Announce Type: new Abstract: Flow matching and diffusion models enable conditional generation across domains ranging from images to proteins, with recent extensions to out-of-distri

researcharxiv-cs-lg
11 Jun 2026
Model Releases

Grounding Computer Use Agents on Human Demonstrations

DGX agent

arXiv:2511.07332v2 Announce Type: replace-cross Abstract: Building reliable computer-use agents requires grounding: accurately connecting natural language instructions to the correct on-screen element

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Hey Chat, Can You Teach Me? Structuring Socratic Dialogue for Human Learning in the Wild

DGX agent

arXiv:2606.11744v1 Announce Type: cross Abstract: Large language models are now widely used for everyday learning, but the underlying interactions are typically unstructured chats rather than followin

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

How Auxiliary Reasoning Unleashes GUI Grounding in VLMs

DGX agent

arXiv:2509.11548v2 Announce Type: replace Abstract: Graphical user interface (GUI) grounding is a fundamental task for building GUI agents. However, general vision-language models (VLMs) struggle with

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Loss Landscape Diagnosis for Gradient-Based Gray-Scott System Inversion: Disentangling the Roles of PINN Components

DGX agent

arXiv:2606.11258v1 Announce Type: new Abstract: Gradient-based inversion of reaction-diffusion systems is typically approached via surrogate models or physics-informed neural networks (PINNs), while t

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

MPC-Patch-Bench: Security-Aware LLM Code Patch for Multi-Party Computation

DGX agent

arXiv:2606.11416v1 Announce Type: cross Abstract: Repository-level benchmarks for evaluating Large Language Model (LLM) code repair on Secure Multi-Party Computation (MPC) software do not yet exist, a

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

OSCS-SupCon: Orthogonal Sigmoid-based Common and Style Supervised Contrastive Learning for Robust Feature Disentanglement

DGX agent

arXiv:2606.11233v1 Announce Type: new Abstract: Supervised Contrastive Learning (SupCon) has achieved strong performance by explicitly modeling pairwise relationships among samples. However, existing

model-releasesarxiv-cs-cv
11 Jun 2026
← Previous
1…517518519520521…1369
Next →