AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

A Survey of Heterogeneous Graph Neural Networks for Cybersecurity Anomaly Detection

DGX agent

arXiv:2510.26307v3 Announce Type: replace-cross Abstract: Anomaly detection is a critical task in cybersecurity, where identifying insider threats, access violations, and coordinated attacks is essent

model-releasesarxiv-cs-lg
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Active Flow Expansion for Out-of-Distribution Discovery: from Theory to Molecules

DGX agent

arXiv:2606.08802v1 Announce Type: new Abstract: Standard flow and diffusion pre-training matches the distribution of available data (e.g., molecules), which often covers only a small fraction of the v

local-aiarxiv-cs-lg
9 Jun 2026
Model Releases

AutoMegaKernel: A Statically-Checked Agent Harness for Self-Retargeting Megakernel Synthesis

DGX agent

arXiv:2606.09682v1 Announce Type: new Abstract: AutoMegaKernel (AMK) compiles a HuggingFace Llama-family model into a single persistent cooperative CUDA kernel that runs the whole forward pass in one

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Causal Longitudinal Prior-Fitted Networks for Counterfactual Outcome Prediction

DGX agent

arXiv:2606.05797v2 Announce Type: replace Abstract: Longitudinal treatment decisions from multivariate time-series data require predicting potential outcomes under future treatment sequences in the pr

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage Sites in China

DGX agent

arXiv:2606.08959v1 Announce Type: new Abstract: We introduce ChinaHeritaQA, a multimodal benchmark dataset for evaluating the cultural reasoning abilities of vision-language models (VLMs) on UNESCO Wo

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

End-to-End Training for Discrete Token LLM based TTS System

DGX agent

arXiv:2606.09234v1 Announce Type: cross Abstract: Recent state-of-the-art (SOTA) text-to-speech (TTS) systems typically adopt a cascaded pipeline consisting of a speech tokenizer, an autoregressive la

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ERBench: A Benchmark and Testsuite for Equation Discovery Algorithms

DGX agent

arXiv:2606.09276v1 Announce Type: new Abstract: Equation discovery aims to automate the discovery of scientific models in the form of mathematical equations from data. Technically, equation discovery

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

DGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GD-MIL: Grade-Disentangled Multiple Instance Learning for Multimodal Biochemical Recurrence Prediction in Prostate Cancer

DGX agent

arXiv:2606.09453v1 Announce Type: new Abstract: Biochemical recurrence (BCR) after radical prostatectomy is a critical endpoint in prostate cancer, yet risk stratification relies almost entirely on va

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Generalization in Nonlinear Least Squares via Learned Feature Geometry

DGX agent

arXiv:2606.08799v1 Announce Type: cross Abstract: We study the generalization of ridge-regularized nonlinear least-squares models via on-average algorithmic stability, deriving error bounds for local

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

GIScholarBench: Benchmarking LLM Overconfidence in GIS Research

DGX agent

arXiv:2606.08036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in academic research workflows, but scholarly tasks require high factual precision and therefore ex

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GRPO Does Not Close the Multi-Agent Coordination Gap

DGX agent

arXiv:2606.07845v1 Announce Type: cross Abstract: We measure how well current large language models coordinate as multiple agents sharing a common resource, using the dining philosophers problem as a

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Guided Discovery of New Behaviors using Diffusion Policies

DGX agent

arXiv:2606.08743v1 Announce Type: new Abstract: Diffusion models have become a powerful tool for generative modeling in robotics, with diffusion policies excelling at modeling multimodal action-trajec

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

Hybrid Robustness Verification for Spatio-Temporal Neural Networks

DGX agent

arXiv:2606.09746v1 Announce Type: cross Abstract: With AI increasingly deployed in safety-critical systems, providing formal robustness guarantees for the underlying models is essential. Existing veri

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Language-based Trial and Error Falls Behind in the Era of Experience

DGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Learning from Human Driving: A Human-in-the-Loop Online Behavior Cloning Framework for Autonomous Driving

DGX agent

arXiv:2606.08170v1 Announce Type: new Abstract: With the evolution of large foundation models (LFMs), data-driven autonomous driving has made significant strides. However, existing paradigms still fac

model-releasesarxiv-cs-ro
9 Jun 2026
Local Ai

Mean Teacher based SSL Framework for Indoor Localization Using Wi-Fi RSSI Fingerprinting

DGX agent

arXiv:2407.13303v2 Announce Type: replace Abstract: Conventional large-scale indoor localization based on Wi-Fi RSSI fingerprinting faces issues of time-consuming and labor-intensive labeled data coll

local-aiarxiv-cs-lg
9 Jun 2026
Applications

MedicalRec: Medical recommender system for image classification without retraining

DGX agent

arXiv:2606.07553v1 Announce Type: cross Abstract: The emergence of machine learning and deep learning has revolutionized the efficiency of diagnostic, therapeutic, and administrative systems in health

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

MMR-GRPO: Accelerating GRPO-Style Training through Diversity-Aware Reward Reweighting

DGX agent

arXiv:2601.09085v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has become a standard approach for training mathematical reasoning models; however, its reliance on

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs

DGX agent

arXiv:2606.09411v1 Announce Type: cross Abstract: Large language models can be fine-tuned to encode prompt-borne secrets into fluent, seemingly benign outputs. This creates a steganographic exfiltrati

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

OmniCap-IF: Benchmarking and Improving Instruction Following Abilities for Omni-Video Captioning

DGX agent

arXiv:2606.08572v1 Announce Type: new Abstract: While Omni-modal Large Language Models (OLLMs) have demonstrated impressive capabilities in jointly processing audio and visual streams, their ability t

model-releasesarxiv-cs-cv
9 Jun 2026
Research

Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation

DGX agent

arXiv:2606.08011v1 Announce Type: cross Abstract: Although directly prompting off-the-shelf Large Language Models (LLMs) to generate meaning-preserving source rewrites can effectively enhance Machine

researcharxiv-cs-ai
9 Jun 2026
Model Releases

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

DGX agent

arXiv:2606.09064v1 Announce Type: cross Abstract: Recent advances in Video Large Language Models (Video-LLMs) have enabled performance on long-video understanding tasks. However, existing methods stil

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation

DGX agent

arXiv:2606.08278v1 Announce Type: new Abstract: Humanoid foundation models are advancing faster than we can evaluate them. While real-world testing is expensive and difficult to reproduce, existing si

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks

DGX agent

arXiv:2606.09669v1 Announce Type: new Abstract: Spatial reasoning is a foundational capability for multimodal large language models (MLLMs) to perceive and operate within the physical world. However,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Still: Amortized KV Cache Compaction in a Single Forward Pass

DGX agent

arXiv:2606.07878v1 Announce Type: new Abstract: The KV cache is the memory bottleneck of long-horizon language model deployment. Practically, a deployable compactor must be lightweight enough to call

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Trajectory-Refined Distillation

DGX agent

arXiv:2606.08432v1 Announce Type: new Abstract: On-policy distillation (OPD) has become a central post-training tool for large language models (LLMs), providing dense per-token teacher supervision alo

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Vision-Language Asymmetry in Bistable Image Captioning

DGX agent

arXiv:2606.08031v1 Announce Type: new Abstract: Wittgenstein's duck-rabbit poses a question for vision-language models: when a model captions an ambiguous image, where in the model is the commitment t

safetyarxiv-cs-cv
9 Jun 2026
Research

What's the Point? Spatial Grammar & Index Resolution for Sign Language Processing

DGX agent

arXiv:2606.08056v1 Announce Type: cross Abstract: Sign language models are predominantly trained with gloss-sequence or text supervision, thereby under-modeling non-lexical and productive construction

researcharxiv-cs-ai
9 Jun 2026
Model Releases

A Comprehensive Anatomy of Human and DeepSeek-R1 LLM Mathematical Reasoning

DGX agent

arXiv:2606.07410v1 Announce Type: cross Abstract: The emergence of 'Aha moments' in large language models, particularly DeepSeek-R1-0120, has raised the question of whether these systems genuinely rea

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

A Held-Out Transition-Pair Falsifier for Long-Horizon Non-Abelian State Tracking

DGX agent

arXiv:2606.07254v1 Announce Type: new Abstract: State tracking exposes a sharp limitation of sequence models: the relevant signal is often not a summary of observed tokens, but an ordered latent state

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

ADAGE: Active Defenses Against GNN Extraction

DGX agent

arXiv:2503.00065v4 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) achieve high performance in various real-world applications, such as drug discovery, traffic states prediction, a

model-releasesarxiv-cs-lg
8 Jun 2026
Tutorials

ARAPDiffusion: ARAP Regularization for Diffusion-Based Deformable Shape Space Learning

DGX agent

arXiv:2606.06887v1 Announce Type: new Abstract: This paper introduces ARAPDiffusion, a latent diffusion model to learn the underlying continuous shape space of a deformation shape collection. The key

tutorialsarxiv-cs-cv
8 Jun 2026
Model Releases

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights

DGX agent

arXiv:2606.07020v1 Announce Type: new Abstract: Multilingual and multicultural benchmarks now cover dozens of languages and model families, but the resulting score landscapes remain metric-rich and in

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

NTILC: Neural Tool Invocation via Learned Compression

DGX agent

arXiv:2606.06566v1 Announce Type: cross Abstract: Agentic tool-calling language models depend on large registries of callable APIs, functions, and local actions. Placing full tool specifications direc

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Scenarios

DGX agent

arXiv:2606.06959v1 Announce Type: cross Abstract: Hallucination detection is essential for the reliable deployment of large language models (LLMs). However, existing evaluations face two core challeng

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Product units in gated recurrent units improve nuclear-mass prediction

DGX agent

arXiv:2606.06866v1 Announce Type: new Abstract: The prediction of masses of atomic nuclei using machine learning can complement theoretical models and advance the exploration of poorly known domains o

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

DGX agent

arXiv:2510.17004v2 Announce Type: replace-cross Abstract: Purpose: To develop and evaluate a multi-agent framework (ReclAIm) for automated monitoring, detection, and correction of performance decline

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling

DGX agent

arXiv:2606.07404v1 Announce Type: new Abstract: This paper reports on training a hundred-billion-parameter sparse mixture of experts on a single eight-GPU node, end to end. LightningLM 0.1V is a recur

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

RhinoVLA Technical Report

DGX agent

arXiv:2606.07383v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for robotic manipulation, but real-time deployment on edge hardware remains challengin

model-releasesarxiv-cs-lg
8 Jun 2026
Research

Robustly estimating heterogeneity in factorial data using Rashomon Partitions

DGX agent

arXiv:2404.02141v5 Announce Type: replace-cross Abstract: In both observational data and randomized control trials, researchers select statistical models to articulate how the outcome of interest vari

researcharxiv-cs-lg
8 Jun 2026
Applications

Sparsely gated tiny linear experts

DGX agent

arXiv:2606.07414v1 Announce Type: new Abstract: Sparsity allows scaling model parameters without proportionally increasing computational cost. While mixture of experts (MoE) models are made increasing

applicationsarxiv-cs-lg
8 Jun 2026
Model Releases

The Post-GCN Decade Revisited: Curvature-Stratified Evaluation of Relational Learning

DGX agent

arXiv:2606.06397v2 Announce Type: replace Abstract: Current evaluation practices in relational learning rely heavily on flat leaderboards that average performance across heterogeneous datasets, implic

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Trading Engagement for Sustainability: Carbon-Aware Re-ranking for E-commerce Recommendations

DGX agent

arXiv:2606.04550v1 Announce Type: cross Abstract: E-commerce recommender systems strongly influence which products users consider and purchase, yet sustainability signals such as Product Carbon Footpr

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Can LLMs Write Correct TLA+ Specifications? Evaluating Natural-Language-to-TLA+ Generation

DGX agent

arXiv:2606.05792v1 Announce Type: new Abstract: TLA+ has supported industrial verification at companies such as Amazon and Microsoft, yet writing correct TLA+ specifications from natural language stil

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs

DGX agent

arXiv:2606.05966v1 Announce Type: cross Abstract: Understanding and reasoning about the physical world is the foundation of intelligent behavior, yet state-of-the-art vision-language models (VLMs) sti

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

GITCO: Gated Inference-Time Context Optimization in TSFMs

DGX agent

arXiv:2606.05332v1 Announce Type: new Abstract: Patch-based Time Series Foundation Models (TSFMs) suffer from context poisoning: structurally anomalous patches capture disproportionate attention and s

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Reward-Decomposed Reinforcement Learning for Immersive Video Role-Playing

DGX agent

arXiv:2605.04733v2 Announce Type: replace Abstract: Text-based role-playing models can imitate character styles, but often fail to capture scene atmosphere and evolving tension, which are crucial for

model-releasesarxiv-cs-ai
6 Jun 2026
← Previous
1…315316317318319…1065
Next →