AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,766 results
Model Releases

SPACE: Source-free Proxy Anchor Concept Erasure for MLLMs

DGX agent

arXiv:2606.09868v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) face growing privacy risks and regulatory constraints, machine unlearning (MU) has emerged as a crucial so

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Research

VFUSE: Virulent Feature Understanding with Sparse autoEncoders

DGX agent

arXiv:2606.10080v1 Announce Type: cross Abstract: Generative models have shown remarkable progress in a variety of domains such as protein design, but such power enables the opaque generation of hazar

researcharxiv-cs-ai
10 Jun 2026
Model Releases

A Survey of Heterogeneous Graph Neural Networks for Cybersecurity Anomaly Detection

DGX agent

arXiv:2510.26307v3 Announce Type: replace-cross Abstract: Anomaly detection is a critical task in cybersecurity, where identifying insider threats, access violations, and coordinated attacks is essent

model-releasesarxiv-cs-lg
9 Jun 2026
Local Ai

Active Flow Expansion for Out-of-Distribution Discovery: from Theory to Molecules

DGX agent

arXiv:2606.08802v1 Announce Type: new Abstract: Standard flow and diffusion pre-training matches the distribution of available data (e.g., molecules), which often covers only a small fraction of the v

local-aiarxiv-cs-lg
9 Jun 2026
Model Releases

AutoMegaKernel: A Statically-Checked Agent Harness for Self-Retargeting Megakernel Synthesis

DGX agent

arXiv:2606.09682v1 Announce Type: new Abstract: AutoMegaKernel (AMK) compiles a HuggingFace Llama-family model into a single persistent cooperative CUDA kernel that runs the whole forward pass in one

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Causal Longitudinal Prior-Fitted Networks for Counterfactual Outcome Prediction

DGX agent

arXiv:2606.05797v2 Announce Type: replace Abstract: Longitudinal treatment decisions from multivariate time-series data require predicting potential outcomes under future treatment sequences in the pr

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage Sites in China

DGX agent

arXiv:2606.08959v1 Announce Type: new Abstract: We introduce ChinaHeritaQA, a multimodal benchmark dataset for evaluating the cultural reasoning abilities of vision-language models (VLMs) on UNESCO Wo

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

End-to-End Training for Discrete Token LLM based TTS System

DGX agent

arXiv:2606.09234v1 Announce Type: cross Abstract: Recent state-of-the-art (SOTA) text-to-speech (TTS) systems typically adopt a cascaded pipeline consisting of a speech tokenizer, an autoregressive la

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ERBench: A Benchmark and Testsuite for Equation Discovery Algorithms

DGX agent

arXiv:2606.09276v1 Announce Type: new Abstract: Equation discovery aims to automate the discovery of scientific models in the form of mathematical equations from data. Technically, equation discovery

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

DGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GD-MIL: Grade-Disentangled Multiple Instance Learning for Multimodal Biochemical Recurrence Prediction in Prostate Cancer

DGX agent

arXiv:2606.09453v1 Announce Type: new Abstract: Biochemical recurrence (BCR) after radical prostatectomy is a critical endpoint in prostate cancer, yet risk stratification relies almost entirely on va

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Generalization in Nonlinear Least Squares via Learned Feature Geometry

DGX agent

arXiv:2606.08799v1 Announce Type: cross Abstract: We study the generalization of ridge-regularized nonlinear least-squares models via on-average algorithmic stability, deriving error bounds for local

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

GIScholarBench: Benchmarking LLM Overconfidence in GIS Research

DGX agent

arXiv:2606.08036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in academic research workflows, but scholarly tasks require high factual precision and therefore ex

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GRPO Does Not Close the Multi-Agent Coordination Gap

DGX agent

arXiv:2606.07845v1 Announce Type: cross Abstract: We measure how well current large language models coordinate as multiple agents sharing a common resource, using the dining philosophers problem as a

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Guided Discovery of New Behaviors using Diffusion Policies

DGX agent

arXiv:2606.08743v1 Announce Type: new Abstract: Diffusion models have become a powerful tool for generative modeling in robotics, with diffusion policies excelling at modeling multimodal action-trajec

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

Hybrid Robustness Verification for Spatio-Temporal Neural Networks

DGX agent

arXiv:2606.09746v1 Announce Type: cross Abstract: With AI increasingly deployed in safety-critical systems, providing formal robustness guarantees for the underlying models is essential. Existing veri

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Language-based Trial and Error Falls Behind in the Era of Experience

DGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Learning from Human Driving: A Human-in-the-Loop Online Behavior Cloning Framework for Autonomous Driving

DGX agent

arXiv:2606.08170v1 Announce Type: new Abstract: With the evolution of large foundation models (LFMs), data-driven autonomous driving has made significant strides. However, existing paradigms still fac

model-releasesarxiv-cs-ro
9 Jun 2026
Local Ai

Mean Teacher based SSL Framework for Indoor Localization Using Wi-Fi RSSI Fingerprinting

DGX agent

arXiv:2407.13303v2 Announce Type: replace Abstract: Conventional large-scale indoor localization based on Wi-Fi RSSI fingerprinting faces issues of time-consuming and labor-intensive labeled data coll

local-aiarxiv-cs-lg
9 Jun 2026
Applications

MedicalRec: Medical recommender system for image classification without retraining

DGX agent

arXiv:2606.07553v1 Announce Type: cross Abstract: The emergence of machine learning and deep learning has revolutionized the efficiency of diagnostic, therapeutic, and administrative systems in health

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

MMR-GRPO: Accelerating GRPO-Style Training through Diversity-Aware Reward Reweighting

DGX agent

arXiv:2601.09085v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has become a standard approach for training mathematical reasoning models; however, its reliance on

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs

DGX agent

arXiv:2606.09411v1 Announce Type: cross Abstract: Large language models can be fine-tuned to encode prompt-borne secrets into fluent, seemingly benign outputs. This creates a steganographic exfiltrati

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

OmniCap-IF: Benchmarking and Improving Instruction Following Abilities for Omni-Video Captioning

DGX agent

arXiv:2606.08572v1 Announce Type: new Abstract: While Omni-modal Large Language Models (OLLMs) have demonstrated impressive capabilities in jointly processing audio and visual streams, their ability t

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Report: GKE Inference Gateway delivers up to 92% faster AI responses

DGX agent

As generative AI moves from experimental pilots to massive production environments, the efficiency of your infrastructure becomes the ultimate differentiator. One way to get the most out of it and min

model-releasesgoogle-cloud-ai
9 Jun 2026
Research

Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation

DGX agent

arXiv:2606.08011v1 Announce Type: cross Abstract: Although directly prompting off-the-shelf Large Language Models (LLMs) to generate meaning-preserving source rewrites can effectively enhance Machine

researcharxiv-cs-ai
9 Jun 2026
Model Releases

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

DGX agent

arXiv:2606.09064v1 Announce Type: cross Abstract: Recent advances in Video Large Language Models (Video-LLMs) have enabled performance on long-video understanding tasks. However, existing methods stil

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation

DGX agent

arXiv:2606.08278v1 Announce Type: new Abstract: Humanoid foundation models are advancing faster than we can evaluate them. While real-world testing is expensive and difficult to reproduce, existing si

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks

DGX agent

arXiv:2606.09669v1 Announce Type: new Abstract: Spatial reasoning is a foundational capability for multimodal large language models (MLLMs) to perceive and operate within the physical world. However,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Still: Amortized KV Cache Compaction in a Single Forward Pass

DGX agent

arXiv:2606.07878v1 Announce Type: new Abstract: The KV cache is the memory bottleneck of long-horizon language model deployment. Practically, a deployable compactor must be lightweight enough to call

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Trajectory-Refined Distillation

DGX agent

arXiv:2606.08432v1 Announce Type: new Abstract: On-policy distillation (OPD) has become a central post-training tool for large language models (LLMs), providing dense per-token teacher supervision alo

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Vision-Language Asymmetry in Bistable Image Captioning

DGX agent

arXiv:2606.08031v1 Announce Type: new Abstract: Wittgenstein's duck-rabbit poses a question for vision-language models: when a model captions an ambiguous image, where in the model is the commitment t

safetyarxiv-cs-cv
9 Jun 2026
Research

What's the Point? Spatial Grammar & Index Resolution for Sign Language Processing

DGX agent

arXiv:2606.08056v1 Announce Type: cross Abstract: Sign language models are predominantly trained with gloss-sequence or text supervision, thereby under-modeling non-lexical and productive construction

researcharxiv-cs-ai
9 Jun 2026
Model Releases

XCR-Bench: Benchmarking Cross-Cultural Reasoning in LLMs via Culture-Specific Items and Hall's Triad

DGX agent

arXiv:2601.14063v2 Announce Type: replace-cross Abstract: Cross-cultural competence in large language models (LLMs) requires understanding and adapting Culture-Specific Items (CSIs) across varying cul

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A Comprehensive Anatomy of Human and DeepSeek-R1 LLM Mathematical Reasoning

DGX agent

arXiv:2606.07410v1 Announce Type: cross Abstract: The emergence of 'Aha moments' in large language models, particularly DeepSeek-R1-0120, has raised the question of whether these systems genuinely rea

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

A Held-Out Transition-Pair Falsifier for Long-Horizon Non-Abelian State Tracking

DGX agent

arXiv:2606.07254v1 Announce Type: new Abstract: State tracking exposes a sharp limitation of sequence models: the relevant signal is often not a summary of observed tokens, but an ordered latent state

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

ADAGE: Active Defenses Against GNN Extraction

DGX agent

arXiv:2503.00065v4 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) achieve high performance in various real-world applications, such as drug discovery, traffic states prediction, a

model-releasesarxiv-cs-lg
8 Jun 2026
Tutorials

ARAPDiffusion: ARAP Regularization for Diffusion-Based Deformable Shape Space Learning

DGX agent

arXiv:2606.06887v1 Announce Type: new Abstract: This paper introduces ARAPDiffusion, a latent diffusion model to learn the underlying continuous shape space of a deformation shape collection. The key

tutorialsarxiv-cs-cv
8 Jun 2026
Model Releases

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights

DGX agent

arXiv:2606.07020v1 Announce Type: new Abstract: Multilingual and multicultural benchmarks now cover dozens of languages and model families, but the resulting score landscapes remain metric-rich and in

model-releasesarxiv-cs-cl
8 Jun 2026
Industry

Nex-N2-Pro running locally https://huggingface.co/nex-agi/Nex-N2-Pro

DGX agent

Nex-N2-Pro is a model available on Hugging Face that can be run locally, enabling users to execute the model on their own infrastructure rather than relying on cloud services. The model is distributed

industryclem-delangue--x
8 Jun 2026
Model Releases

NTILC: Neural Tool Invocation via Learned Compression

DGX agent

arXiv:2606.06566v1 Announce Type: cross Abstract: Agentic tool-calling language models depend on large registries of callable APIs, functions, and local actions. Placing full tool specifications direc

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Scenarios

DGX agent

arXiv:2606.06959v1 Announce Type: cross Abstract: Hallucination detection is essential for the reliable deployment of large language models (LLMs). However, existing evaluations face two core challeng

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Product units in gated recurrent units improve nuclear-mass prediction

DGX agent

arXiv:2606.06866v1 Announce Type: new Abstract: The prediction of masses of atomic nuclei using machine learning can complement theoretical models and advance the exploration of poorly known domains o

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

DGX agent

arXiv:2510.17004v2 Announce Type: replace-cross Abstract: Purpose: To develop and evaluate a multi-agent framework (ReclAIm) for automated monitoring, detection, and correction of performance decline

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling

DGX agent

arXiv:2606.07404v1 Announce Type: new Abstract: This paper reports on training a hundred-billion-parameter sparse mixture of experts on a single eight-GPU node, end to end. LightningLM 0.1V is a recur

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

RhinoVLA Technical Report

DGX agent

arXiv:2606.07383v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for robotic manipulation, but real-time deployment on edge hardware remains challengin

model-releasesarxiv-cs-lg
8 Jun 2026
Research

Robustly estimating heterogeneity in factorial data using Rashomon Partitions

DGX agent

arXiv:2404.02141v5 Announce Type: replace-cross Abstract: In both observational data and randomized control trials, researchers select statistical models to articulate how the outcome of interest vari

researcharxiv-cs-lg
8 Jun 2026
Applications

Sparsely gated tiny linear experts

DGX agent

arXiv:2606.07414v1 Announce Type: new Abstract: Sparsity allows scaling model parameters without proportionally increasing computational cost. While mixture of experts (MoE) models are made increasing

applicationsarxiv-cs-lg
8 Jun 2026
← Previous
1…407408409410411…1371
Next →