AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
Model Releases

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

DGX agent

arXiv:2606.09064v1 Announce Type: cross Abstract: Recent advances in Video Large Language Models (Video-LLMs) have enabled performance on long-video understanding tasks. However, existing methods stil

model-releasesarxiv-cs-ai
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SIMPLE: Simulation-Based Policy Learning and Evaluation for Humanoid Loco-manipulation

DGX agent

arXiv:2606.08278v1 Announce Type: new Abstract: Humanoid foundation models are advancing faster than we can evaluate them. While real-world testing is expensive and difficult to reproduce, existing si

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks

DGX agent

arXiv:2606.09669v1 Announce Type: new Abstract: Spatial reasoning is a foundational capability for multimodal large language models (MLLMs) to perceive and operate within the physical world. However,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Still: Amortized KV Cache Compaction in a Single Forward Pass

DGX agent

arXiv:2606.07878v1 Announce Type: new Abstract: The KV cache is the memory bottleneck of long-horizon language model deployment. Practically, a deployable compactor must be lightweight enough to call

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Trajectory-Refined Distillation

DGX agent

arXiv:2606.08432v1 Announce Type: new Abstract: On-policy distillation (OPD) has become a central post-training tool for large language models (LLMs), providing dense per-token teacher supervision alo

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Vision-Language Asymmetry in Bistable Image Captioning

DGX agent

arXiv:2606.08031v1 Announce Type: new Abstract: Wittgenstein's duck-rabbit poses a question for vision-language models: when a model captions an ambiguous image, where in the model is the commitment t

safetyarxiv-cs-cv
9 Jun 2026
Research

What's the Point? Spatial Grammar & Index Resolution for Sign Language Processing

DGX agent

arXiv:2606.08056v1 Announce Type: cross Abstract: Sign language models are predominantly trained with gloss-sequence or text supervision, thereby under-modeling non-lexical and productive construction

researcharxiv-cs-ai
9 Jun 2026
Model Releases

XCR-Bench: Benchmarking Cross-Cultural Reasoning in LLMs via Culture-Specific Items and Hall's Triad

DGX agent

arXiv:2601.14063v2 Announce Type: replace-cross Abstract: Cross-cultural competence in large language models (LLMs) requires understanding and adapting Culture-Specific Items (CSIs) across varying cul

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A Comprehensive Anatomy of Human and DeepSeek-R1 LLM Mathematical Reasoning

DGX agent

arXiv:2606.07410v1 Announce Type: cross Abstract: The emergence of 'Aha moments' in large language models, particularly DeepSeek-R1-0120, has raised the question of whether these systems genuinely rea

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

A Held-Out Transition-Pair Falsifier for Long-Horizon Non-Abelian State Tracking

DGX agent

arXiv:2606.07254v1 Announce Type: new Abstract: State tracking exposes a sharp limitation of sequence models: the relevant signal is often not a summary of observed tokens, but an ordered latent state

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

ADAGE: Active Defenses Against GNN Extraction

DGX agent

arXiv:2503.00065v4 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) achieve high performance in various real-world applications, such as drug discovery, traffic states prediction, a

model-releasesarxiv-cs-lg
8 Jun 2026
Tutorials

ARAPDiffusion: ARAP Regularization for Diffusion-Based Deformable Shape Space Learning

DGX agent

arXiv:2606.06887v1 Announce Type: new Abstract: This paper introduces ARAPDiffusion, a latent diffusion model to learn the underlying continuous shape space of a deformation shape collection. The key

tutorialsarxiv-cs-cv
8 Jun 2026
Model Releases

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights

DGX agent

arXiv:2606.07020v1 Announce Type: new Abstract: Multilingual and multicultural benchmarks now cover dozens of languages and model families, but the resulting score landscapes remain metric-rich and in

model-releasesarxiv-cs-cl
8 Jun 2026
Industry

Nex-N2-Pro running locally https://huggingface.co/nex-agi/Nex-N2-Pro

DGX agent

Nex-N2-Pro is a model available on Hugging Face that can be run locally, enabling users to execute the model on their own infrastructure rather than relying on cloud services. The model is distributed

industryclem-delangue--x
8 Jun 2026
Model Releases

NTILC: Neural Tool Invocation via Learned Compression

DGX agent

arXiv:2606.06566v1 Announce Type: cross Abstract: Agentic tool-calling language models depend on large registries of callable APIs, functions, and local actions. Placing full tool specifications direc

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Scenarios

DGX agent

arXiv:2606.06959v1 Announce Type: cross Abstract: Hallucination detection is essential for the reliable deployment of large language models (LLMs). However, existing evaluations face two core challeng

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Product units in gated recurrent units improve nuclear-mass prediction

DGX agent

arXiv:2606.06866v1 Announce Type: new Abstract: The prediction of masses of atomic nuclei using machine learning can complement theoretical models and advance the exploration of poorly known domains o

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

DGX agent

arXiv:2510.17004v2 Announce Type: replace-cross Abstract: Purpose: To develop and evaluate a multi-agent framework (ReclAIm) for automated monitoring, detection, and correction of performance decline

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling

DGX agent

arXiv:2606.07404v1 Announce Type: new Abstract: This paper reports on training a hundred-billion-parameter sparse mixture of experts on a single eight-GPU node, end to end. LightningLM 0.1V is a recur

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

RhinoVLA Technical Report

DGX agent

arXiv:2606.07383v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for robotic manipulation, but real-time deployment on edge hardware remains challengin

model-releasesarxiv-cs-lg
8 Jun 2026
Research

Robustly estimating heterogeneity in factorial data using Rashomon Partitions

DGX agent

arXiv:2404.02141v5 Announce Type: replace-cross Abstract: In both observational data and randomized control trials, researchers select statistical models to articulate how the outcome of interest vari

researcharxiv-cs-lg
8 Jun 2026
Applications

Sparsely gated tiny linear experts

DGX agent

arXiv:2606.07414v1 Announce Type: new Abstract: Sparsity allows scaling model parameters without proportionally increasing computational cost. While mixture of experts (MoE) models are made increasing

applicationsarxiv-cs-lg
8 Jun 2026
Model Releases

The Post-GCN Decade Revisited: Curvature-Stratified Evaluation of Relational Learning

DGX agent

arXiv:2606.06397v2 Announce Type: replace Abstract: Current evaluation practices in relational learning rely heavily on flat leaderboards that average performance across heterogeneous datasets, implic

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Trading Engagement for Sustainability: Carbon-Aware Re-ranking for E-commerce Recommendations

DGX agent

arXiv:2606.04550v1 Announce Type: cross Abstract: E-commerce recommender systems strongly influence which products users consider and purchase, yet sustainability signals such as Product Carbon Footpr

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Can LLMs Write Correct TLA+ Specifications? Evaluating Natural-Language-to-TLA+ Generation

DGX agent

arXiv:2606.05792v1 Announce Type: new Abstract: TLA+ has supported industrial verification at companies such as Amazon and Microsoft, yet writing correct TLA+ specifications from natural language stil

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs

DGX agent

arXiv:2606.05966v1 Announce Type: cross Abstract: Understanding and reasoning about the physical world is the foundation of intelligent behavior, yet state-of-the-art vision-language models (VLMs) sti

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

GITCO: Gated Inference-Time Context Optimization in TSFMs

DGX agent

arXiv:2606.05332v1 Announce Type: new Abstract: Patch-based Time Series Foundation Models (TSFMs) suffer from context poisoning: structurally anomalous patches capture disproportionate attention and s

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Reward-Decomposed Reinforcement Learning for Immersive Video Role-Playing

DGX agent

arXiv:2605.04733v2 Announce Type: replace Abstract: Text-based role-playing models can imitate character styles, but often fail to capture scene atmosphere and evolving tension, which are crucial for

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Before the week ends, let's acknowledge one of the most INSANE week ever for open AI, with 25+ notable open-weight drops across every modali…

DGX agent

Before the week ends, let's acknowledge one of the most INSANE week ever for open AI, with 25+ notable open-weight drops across every modality: 🧠 LLMs → NVIDIA Nemotron 3 Ultra: 550B hybrid Mamba-MoE,

model-releasesclem-delangue--x
5 Jun 2026
Model Releases

Can LLMs Be Constrained to the Past? Improving Knowledge Cutoff through Recall-Based Prompting

DGX agent

arXiv:2606.05804v1 Announce Type: new Abstract: Prompted knowledge cutoff instructs a large language model (LLM) to act as if information beyond a specified cutoff date were unavailable. However, prio

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Compress-Distill: Reasoning Trace Compression for Efficient Knowledge Distillation

DGX agent

arXiv:2606.05988v1 Announce Type: cross Abstract: Reasoning models produce long chain-of-thought traces that are costly to distill and encourage verbose student outputs. We study post-hoc compression

model-releasesarxiv-cs-cl
5 Jun 2026
Research

From Scoring to Explanations: Evaluating SHAP and LLM Rationales for Rubric-based Teaching Quality Assessment

DGX agent

arXiv:2606.05180v1 Announce Type: new Abstract: Automated scoring models are increasingly used to assign rubric-based quality ratings to complex language performances, including classroom transcripts,

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Harmonious Parameter Adaptation in Continual Visual Instruction Tuning for Safety-Aligned MLLMs

DGX agent

arXiv:2511.20158v2 Announce Type: replace Abstract: While continual visual instruction tuning (CVIT) has shown promise in adapting multimodal large language models (MLLMs), existing studies predominan

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

Improving Answer Extraction in Context-based Question Answering Systems Using LLMs

DGX agent

arXiv:2606.06197v1 Announce Type: new Abstract: Question answering (QA) systems have achieved notable progress with the advent of large language models (LLMs). However, they still face challenges in a

model-releasesarxiv-cs-cl
5 Jun 2026
Research

LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memorization in LLMs

DGX agent

arXiv:2606.06286v1 Announce Type: new Abstract: Large language models can reproduce training data, but existing memorization evaluations mostly measure whether models can be forced to do so, rather th

researcharxiv-cs-cl
5 Jun 2026
Model Releases

LongSpace: Exploring Long-Horizon Spatial Memory from Perception to Recall in Video

DGX agent

arXiv:2606.05677v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have advanced image and video understanding and can increasingly handle longer visual inputs. Long-horizon ta

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

PEFT of SLM for Telecommunications Customer Support: A Comparative Study of LoRA Configurations with Energy Consumption Analysis

DGX agent

arXiv:2606.05176v1 Announce Type: new Abstract: While large language models (LLMs) show strong performance in natural language understanding and generation, their evaluation and adaptation to domain-s

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

ReasoningFlow: Discourse Structures for Understanding LLM Reasoning Traces

DGX agent

arXiv:2606.05402v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce reasoning traces with non-linear structures, such as backtracking and self-correction, that complicate the evaluat

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

TensorBench: Benchmarking Coding Agents on a Compiler-Based Tensor Framework

DGX agent

arXiv:2606.05570v1 Announce Type: new Abstract: Repository-level coding benchmarks face a trade-off between task difficulty and evaluation reliability: tasks that challenge frontier models often invol

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators

DGX agent

arXiv:2606.06476v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have shown strong visual reasoning capabilities, their spatial reasoning abilities remain largely constrained to the

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

UltraVR: A Diagnostic Ultra-Resolution Image-VQA Benchmark for Evidence-Grounded Reasoning

DGX agent

arXiv:2606.05576v1 Announce Type: new Abstract: Vision-language models (VLMs) excel on visual question answering and multimodal reasoning benchmarks. Yet their capability on ultra-resolution images -

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

A Cookbook of 3D Vision: Data, Learning Paradigms, and Application

DGX agent

arXiv:2606.04291v1 Announce Type: new Abstract: 3D vision has rapidly evolved, driven by increasingly diverse data representations, learning paradigms, and modeling strategies. Yet the field remains f

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Adaptive Minds: Empowering Agents with LoRA-as-Tools

DGX agent

arXiv:2510.15416v2 Announce Type: replace Abstract: We investigate a framework in which LoRA adapters are treated as callable tools that a base language model can dynamically select and invoke. We hyp

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

An Empirical Audit of Input Encoders for Multi-Channel Signal Transformers

DGX agent

arXiv:2606.04752v1 Announce Type: cross Abstract: Transformers consuming multi-channel scalar signals must embed C simultaneous values into one d_{ext{model}}-dimensional vector per time step. We empi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Bypassing Prompt Guards in Production with Controlled-Release Prompting

DGX agent

arXiv:2510.01529v3 Announce Type: replace Abstract: Ball et al. recently established that prompt filtering for AI alignment faces a fundamental barrier: under standard cryptographic assumptions, no fi

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Can I Take Another Dose? Evaluating LLM Decision-Making Under Temporal Uncertainty in OTC Dosing QA

DGX agent

arXiv:2606.04262v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for everyday health questions, including whether a user can safely take another dose of an over-the

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Enhancing Hallucination Detection through Noise Injection

DGX agent

arXiv:2502.03799v4 Announce Type: replace Abstract: Large Language Models (LLMs) are prone to generating plausible yet incorrect responses, known as hallucinations. Effectively detecting hallucination

researcharxiv-cs-cl
4 Jun 2026
Agents

FALSIFYBENCH: Evaluating Inductive Reasoning in LLMs with Rule Discovery Games

DGX agent

arXiv:2606.04751v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in scientific tasks. Yet whether these systems can effectively engage in for

agentsarxiv-cs-ai
4 Jun 2026
← Previous
1…393394395396397…1326
Next →