AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,387
  • Agents7,863
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,238
  • Local Ai5,173
  • Model Releases25,258
  • Research21,121
  • Safety13,950
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,387
  • Agents7,863
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,238
  • Local Ai5,173
  • Model Releases25,258
  • Research21,121
  • Safety13,950
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,387Total entries
1Added by human
92,386Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,911 results
Safety

DeMix: Debugging Training Data with Mixed Data Error Types by Investigating Influence Vectors

DGX agent

arXiv:2606.11616v1 Announce Type: new Abstract: High-quality training data is essential for the success of machine learning models. However, real-world datasets often contain mixed types of errors ari

safetyarxiv-cs-lg
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners?

DGX agent

arXiv:2606.12402v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as high-level planners for embodied agents, with an emerging strategy of scaling test-time com

applicationsarxiv-cs-ai
11 Jun 2026
Tutorials

FOCUS: DLLMs Know How to Tame Their Compute Bound

DGX agent

arXiv:2601.23278v2 Announce Type: replace-cross Abstract: Diffusion Large Language Models (DLLMs) offer a compelling alternative to Auto-Regressive models, but their deployment is constrained by high

tutorialsarxiv-cs-cl
11 Jun 2026
Model Releases

GraspLLM: Towards Zero-Shot Generalization on Text-Attributed Graphs with LLMs

DGX agent

arXiv:2606.11898v1 Announce Type: new Abstract: Research on Text-Attributed Graphs (TAGs) has gained significant attention recently due to its broad applications across various real-world data scenari

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Human-Guided Agentic AI for Multimodal Clinical Prediction: Lessons from the AgentDS Healthcare Benchmark

DGX agent

arXiv:2602.19502v2 Announce Type: replace Abstract: Agentic AI systems are increasingly capable of autonomous data science workflows, yet clinical prediction tasks demand domain expertise that purely

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Machine-learning clustering of close-in exoplanet populations: links to pebble accretion

DGX agent

arXiv:2606.11737v1 Announce Type: cross Abstract: Close-in exoplanets exhibit a wide range of orbital architectures and physical properties shaped by both formation conditions and migration processes.

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Mind the Perspective: Let's Reason Recursively for Theory of Mind

DGX agent

arXiv:2606.11724v1 Announce Type: new Abstract: Theory of Mind (ToM) reasoning requires inferring agents' beliefs from partial and asymmetric observations, which remains an open challenge for LLMs. Ex

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

MSUE: Multi-Modal Soccer Understanding Expert

DGX agent

arXiv:2606.12106v1 Announce Type: cross Abstract: This paper presents our solution to the 2026 SoccerNet VQA Challenge. We first develop a cost-effective data synthesis pipeline driven by a Vision-Lan

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Multi-Rate Mixture of Experts for Accelerating Liquid Neural Network Training

DGX agent

arXiv:2606.12240v1 Announce Type: cross Abstract: Multivariate time-series data often exhibit complex temporal dependencies, irregular sampling, and heterogeneous dynamics across multiple time scales,

researcharxiv-cs-ai
11 Jun 2026
Model Releases

Multi-View In-Cabin Monitoring System for Public Transport Vehicles

DGX agent

arXiv:2606.11739v1 Announce Type: cross Abstract: We introduce a multi-view in-cabin monitoring dataset for public transportation with synchronized RGB and depth images from four inward-facing cameras

model-releasesarxiv-cs-ai
11 Jun 2026
Research

NSVQ: Mitigating Codebook Collapse by Stabilizing Encoder Drift in Vector Quantization

DGX agent

arXiv:2606.11363v1 Announce Type: new Abstract: Vector quantization is central to modern generative modeling pipelines, but large-codebook VQ models often suffer from codebook collapse. We identify en

researcharxiv-cs-cv
11 Jun 2026
Model Releases

OCSVM-Guided Representation Learning for Unsupervised Anomaly Detection

DGX agent

arXiv:2507.21164v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection (UAD) aims to detect anomalies without labeled data, a necessity in many machine learning applications where an

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

On the Limits of LLM-as-Judge for Scientific Novelty Assessment

DGX agent

arXiv:2606.12071v1 Announce Type: cross Abstract: LLMs are increasingly used to generate and judge scientific ideas. This makes novelty evaluation a central problem. Full idea evaluation is difficult

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Open Materials Generation with Inference-Time Reinforcement Learning

DGX agent

arXiv:2602.00424v2 Announce Type: replace Abstract: Continuous-time generative models for crystalline materials enable inverse materials design by learning to predict stable crystal structures, but in

safetyarxiv-cs-lg
11 Jun 2026
Model Releases

PCS-UQ: Uncertainty Quantification via the Predictability-Computability-Stability Framework

DGX agent

arXiv:2505.08784v2 Announce Type: replace-cross Abstract: As machine learning (ML) enters high-stakes domains, trustworthy uncertainty quantification (UQ) is essential for safety. In this paper we int

model-releasesarxiv-cs-lg
11 Jun 2026
Tutorials

Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!

DGX agent

arXiv:2504.09762v4 Announce Type: replace Abstract: Intermediate token generation (ITG), where a model produces output before the solution, has become a standard method to improve the performance of l

tutorialsarxiv-cs-ai
11 Jun 2026
Research

Projected random forests and conformal prediction of circular data

DGX agent

arXiv:2410.24145v3 Announce Type: replace-cross Abstract: We apply conformal prediction techniques to regression problems with circular responses, producing prediction sets with adaptive arc length an

researcharxiv-cs-lg
11 Jun 2026
Model Releases

ReMoT: Reinforcement Learning with Motion Contrast Triplets

DGX agent

arXiv:2603.00461v3 Announce Type: replace Abstract: We present ReMoT, a unified training paradigm to systematically address the fundamental shortcomings of VLMs in spatio-temporal consistency -- a cri

model-releasesarxiv-cs-cv
11 Jun 2026
Research

The Long Tail, Not the Front Page: Cold-Start Prediction of Crowd Highlight Salience

DGX agent

arXiv:2606.11654v1 Announce Type: cross Abstract: A social highlighter's most useful signal -- which passages a crowd of readers marks -- exists only for documents people have already read. Can the ag

researcharxiv-cs-cl
11 Jun 2026
Model Releases

The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content

DGX agent

arXiv:2606.11198v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems inject external knowledge to improve LLM outputs, yet the format of injected content -- distinct from its

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Time-Conditioned and Multi-Time Survival Prediction from 2D PET/CT Projections in Lung Cancer

DGX agent

arXiv:2606.12140v1 Announce Type: new Abstract: Accurate prediction of overall survival (OS) from positron emission tomography/computed tomography (PET/CT) can support personalized treatment and follo

researcharxiv-cs-cv
11 Jun 2026
Model Releases

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

DGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fabl…

DGX agent

Up until yesterday, our entire MTS team has operated under the philosophy of tokenmaxxing as much as possible on Claude Max plans. With Fable, this may no longer be possible: - One of our team members

model-releasesjerry-liu--x
11 Jun 2026
Tutorials

When is Your LLM Steerable?

DGX agent

arXiv:2606.11599v1 Announce Type: new Abstract: Activation steering offers a lightweight approach to control language models' behavior at inference time, but whether it succeeds or fails heavily depen

tutorialsarxiv-cs-cl
11 Jun 2026
Safety

Alignment Defends LLMs from Property Inference Attacks

DGX agent

arXiv:2606.10217v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly fine-tuned on domain-specific datasets that may contain sensitive, dataset-level properties. Recent work h

safetyarxiv-cs-lg
10 Jun 2026
Safety

AnimaSpark: A Feed-Forward Method for Animating Arbitrary 3D Objects

DGX agent

arXiv:2606.10988v1 Announce Type: new Abstract: While recent advancements in generative AI have substantially accelerated static 3D model creation workflows, the synthesis of category-agnostic 3D anim

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

au-Rec: A Verifiable Benchmark for Agentic Recommender Systems

DGX agent

arXiv:2606.10156v1 Announce Type: cross Abstract: As recommender systems transition toward agentic, multi-turn conversational interfaces, evaluation paradigms have struggled to keep pace. Current benc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Benchmarking and Exploring the Capabilities of LLMs for Attack Investigations

DGX agent

arXiv:2606.10281v1 Announce Type: cross Abstract: This paper presents AuditBench, a new benchmark dataset for evaluating the capabilities of LLMs at investigating security-related system audit logs. W

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Cybersecurity researchers complain that Claude Fable's guardrails are too strict, rejecting 'innocuous tasks' like reading blog posts or performing code reviews (Lorenzo Franceschi-Bicchierai/TechCrunch)

DGX agent

Lorenzo Franceschi-Bicchierai / TechCrunch: Cybersecurity researchers complain that Claude Fable's guardrails are too strict, rejecting “innocuous tasks” like reading blog posts or performing code rev

model-releasestechmeme
10 Jun 2026
Applications

Data-aware Static Analysis: Improving Detection of Semantic Faults in Machine Learning Code Using Data Characteristics

DGX agent

arXiv:2606.09957v1 Announce Type: cross Abstract: Semantic faults specific to the use of machine learning models are a common problem for machine learning developers, causing suboptimal predictions, h

applicationsarxiv-cs-lg
10 Jun 2026
Model Releases

Divide and Cooperate: Role-Decomposed Multi-Agent LLM Training with Cross-Agent Learning Signals

DGX agent

arXiv:2606.10684v1 Announce Type: cross Abstract: Modern language agents which perform multi-step reasoning have shown strong performance in knowledge-intensive question answering. However, existing a

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Dynamic Linear Attention

DGX agent

arXiv:2606.10650v1 Announce Type: cross Abstract: The scalability of Large Language Models (LLMs) to long contexts is fundamentally constrained by the quadratic complexity of standard attention, motiv

researcharxiv-cs-ai
10 Jun 2026
Model Releases

FreshRetailNet-LT: A Stockout-Annotated Censored Demand Dataset for Latent Demand Recovery and Forecasting in Fresh Retail

DGX agent

arXiv:2505.16319v3 Announce Type: replace Abstract: Accurate demand estimation is critical for the retail business in guiding the inventory and pricing policies of perishable products. However, it fac

model-releasesarxiv-cs-lg
10 Jun 2026
Applications

From Senses to Decisions: The Information Flow of Auditory and Visual Perception in Multimodal LLMs

DGX agent

arXiv:2606.10147v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) can listen and see, but how do audio and visual signals actually travel through the network to shape an answer?

applicationsarxiv-cs-ai
10 Jun 2026
Safety

Gradient-Guided Reward Optimization for Inference-time Alignment

DGX agent

arXiv:2606.09635v1 Announce Type: cross Abstract: Ensuring the reliability of Large Language Models (LLMs) under distribution drift requires inference-time adaptation. While inference-time alignment m

safetyarxiv-cs-lg
10 Jun 2026
Model Releases

HISTORY LESSON: In 1968 the US, USSR, UK, France, and China signed the Nuclear Non-Proliferation Treaty, declaring nuclear weapons too dange…

DGX agent

HISTORY LESSON: In 1968 the US, USSR, UK, France, and China signed the Nuclear Non-Proliferation Treaty, declaring nuclear weapons too dangerous for any more countries to build. All five already had t

model-releasesclem-delangue--x
10 Jun 2026
Model Releases

How can we assess human-agent interactions? Case studies in software agent design

DGX agent

arXiv:2510.09801v3 Announce Type: replace Abstract: While benchmarks measure the accuracy of LLM-powered agents, they mostly assume full automation, failing to represent the collaborative nature of re

model-releasesarxiv-cs-ai
10 Jun 2026
Applications

Intelligence layer becomes the enterprise AI control plane for enterprise AI

DGX agent

As enterprises accelerate past AI experimentation into full-scale production, the central challenge has shifted from accessing models to managing the organizational context those models need to act re

applicationssiliconangle
10 Jun 2026
Model Releases

JGRA: Jacobian Geometry Robustness Assessment in NISQ Noise-Aware Quantum Neural Networks

DGX agent

arXiv:2606.09964v1 Announce Type: cross Abstract: The NISQ era places stringent constraints on quantum computation, where noise and decoherence fundamentally limit performance. In classical deep learn

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Latent Guided Sampling for Combinatorial Optimization

DGX agent

arXiv:2506.03672v2 Announce Type: replace-cross Abstract: Combinatorial Optimization problems are widespread in domains such as logistics, manufacturing, and drug discovery, yet their NP-hard nature m

model-releasesarxiv-cs-lg
10 Jun 2026
Safety

Lightweight Latent Reasoning for Narrative Tasks

DGX agent

arXiv:2512.02240v2 Announce Type: replace Abstract: Large language models (LLMs) tackle complex tasks by generating long chains of thought or 'reasoning traces' that act as latent variables in the gen

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

Me, 2024. LLMs will be commodity; (except for Nvdia) profits will be hard to squeeze out. Techbros: Shut up, Gary. GPT-5 is gonna be AGI. To…

DGX agent

Me, 2024. LLMs will be commodity; (except for Nvdia) profits will be hard to squeeze out. Techbros: Shut up, Gary. GPT-5 is gonna be AGI. Today: LLMs are commodity; (except for Nvidia) profits have be

model-releasesgary-marcus--x
10 Jun 2026
Model Releases

Non-Parametric Structural Priors for Geometry Theorem Prediction

DGX agent

arXiv:2603.04852v2 Announce Type: replace Abstract: Multi-step theorem prediction is a central challenge in geometry problem solving. Existing neural-symbolic approaches rely heavily on supervised par

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

OncoTraj: a public benchmark for longitudinal resistance prediction in EGFR-mutant non-small-cell lung cancer on osimertinib

DGX agent

arXiv:2606.11144v1 Announce Type: new Abstract: Resistance to first-line osimertinib in EGFR-mutant non-small-cell lung cancer (NSCLC) is the canonical example of predictable clonal evolution under th

model-releasesarxiv-cs-lg
10 Jun 2026
Research

One Token per Multimodal Evidence: Latent Memory for Resource-Constrained QA

DGX agent

arXiv:2606.10572v1 Announce Type: new Abstract: External memory effectively grounds large language models (LLMs) and vision-language models (VLMs)-based question answering (QA) in relevant multimodal

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Optimization-based Online Conformal Prediction for Multi-step Forecasting

DGX agent

arXiv:2508.13362v2 Announce Type: replace Abstract: Conformal prediction (CP) is well-suited for uncertainty quantification in time series forecasting due to its distribution-free coverage guarantees.

model-releasesarxiv-cs-lg
10 Jun 2026
Research

Optimizing 2D Input Representations and Sub-phase Fusion Strategies for Differential Diagnosis of Asthma and COPD Using CNN- and GRU-Based Networks

DGX agent

arXiv:2606.10972v1 Announce Type: cross Abstract: This study aims to explore the performance of the VAR model in comparison with mel-frequency cepstral coefficient (MFCC) matrices and log-mel spectrog

researcharxiv-cs-ai
10 Jun 2026
Safety

PADD: Path-Aligned Decompression Distillation for Non-Router Teacher to Guide MoE Student Learning

DGX agent

arXiv:2606.10369v1 Announce Type: new Abstract: As large language models (LLMs) continue to scale, it becomes increasingly challenging to grow model capacity under fixed computation budgets. We propos

safetyarxiv-cs-cl
10 Jun 2026
← Previous
1…610611612613614…1394
Next →