AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation

DGX agent

arXiv:2602.02220v2 Announce Type: replace Abstract: Language-conditioned goal navigation (LGN) requires agents to locate user-specified targets without step-by-step guidance. However, existing benchma

model-releasesarxiv-cs-cv
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM Understanding of Paired-Focus Constructions

DGX agent

arXiv:2605.31586v1 Announce Type: cross Abstract: Grasping the semantics of rare constructions (form-meaning pairings) has been shown to be a challenging problem that has currently only been solved by

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Learning-Based Navigation for Indoor Mobile Robots

DGX agent

arXiv:2605.30468v1 Announce Type: new Abstract: This paper presents a learning-based navigation framework for indoor mobile robots. The proposed method combines a supervised neural global planner, tra

model-releasesarxiv-cs-ro
1 Jun 2026
Model Releases

Learning Multi-Agent Coordination via Sheaf-ADMM

DGX agent

arXiv:2605.31005v1 Announce Type: new Abstract: We present a differentiable optimization framework for multi-agent coordination. An input is decomposed into overlapping local views, each processed by

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Learning Randomized Reductions

DGX agent

arXiv:2412.18134v4 Announce Type: replace Abstract: Randomized self-reductions (RSRs) express f(x) using f evaluated at random correlated points, enabling self-correcting programs, instance-hiding pro

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Learning Whom to Trust: Market-Feedback Adaptive Retrieval for Frozen LLMs in Event-Driven Financial RAG

DGX agent

arXiv:2605.31201v1 Announce Type: new Abstract: Financial retrieval-augmented generation (RAG) systems typically rank evidence by textual relevance, but in financial markets the useful evidence source

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

LegSegNet: A Public Deep Learning System for Lower Extremity CT Tissue Segmentation and Quantification

DGX agent

arXiv:2605.30829v1 Announce Type: new Abstract: Lower extremity computed tomography (CT) contains clinically relevant information for body composition analysis, sarcopenia assessment, and musculoskele

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Linear Ordering Problem: Time for a Change

DGX agent

arXiv:2605.31051v1 Announce Type: cross Abstract: The Linear Ordering Problem (LOP) is a fundamental combinatorial optimization problem with important applications in areas such as economics, social c

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

LLM Bias Evaluation: Gender, Racial, and Age Disparities in Occupational and Crime Scenarios

DGX agent

arXiv:2409.14583v4 Announce Type: replace Abstract: LLM bias evaluation is critical as large language models (LLMs) increasingly influence high-stakes decisions. This paper provides a comprehensive as

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

LLMs Without Deep Neural Networks: New Architecture, Benefits and Case Study

DGX agent

arXiv:2605.30385v1 Announce Type: cross Abstract: The purpose of this article is to provide validation to my deep neural network alternative in the context of LLMs. Very recently, there has been a sig

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

DGX agent

arXiv:2605.30434v1 Announce Type: cross Abstract: Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MAAT: Multi-phase Adapter-Aware Targeted Unlearning

DGX agent

arXiv:2605.30514v1 Announce Type: cross Abstract: Machine unlearning evaluation is structurally skewed: Why-type questions, which probe causal and relational knowledge, comprise less than 0.06% of Cou

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

MADS: Model-Aware Diverse Core Set Selection for Instruction Tuning

DGX agent

arXiv:2605.30857v1 Announce Type: new Abstract: Instruction fine-tuning is employed to enhance the instruction-following ability of large language models (LLMs). As the amount of instruction fine-tuni

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

MAVEN: Improving Generalization in Agentic Tool Calling

DGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

DGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Mellum2 Technical Report

DGX agent

arXiv:2605.31268v1 Announce Type: new Abstract: We present Mellum 2, an open-weight 12B-parameter Mixture-of-Experts (MoE) language model with 2.5B active parameters per token. Mellum 2 is a general-p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

DGX agent

arXiv:2605.30571v1 Announce Type: cross Abstract: Physical AI systems, including robots, autonomous vehicles, embodied agents and edge copilots, often run a different inference workload from cloud LLM

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Memory by Design: Probabilistic Sequence Layers

DGX agent

arXiv:2605.31163v1 Announce Type: cross Abstract: We introduce the design-model framework: a way to derive efficient recurrent sequence maps from explicit assumptions about memory. A design model writ

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

MIMO: Multilingual Information Retrieval via Monolingual Objectives

DGX agent

arXiv:2605.31171v1 Announce Type: cross Abstract: Multilingual Information Retrieval (MLIR) reflects real-world search environments in which queries and relevant documents may appear in different lang

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft

DGX agent

arXiv:2605.30931v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and action generation. However, their ability to susta

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials

DGX agent

arXiv:2605.30889v1 Announce Type: cross Abstract: Constructing production-quality machine-learned interatomic potentials (MLIPs) requires balancing accuracy, dynamical stability, and computational thr

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

MosaicLeaks:Privacy Risks in Querying-in-the-Open for Deep Research Agents

DGX agent

arXiv:2605.30727v1 Announce Type: new Abstract: Deep research agents increasingly combine private local documents with external tools like web retrieval, creating a privacy risk: an agent's external q

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

MultiPriv: Benchmarking Individual-Level Privacy Reasoning in Vision-Language Models

DGX agent

arXiv:2511.16940v3 Announce Type: replace Abstract: Modern Vision-Language Models (VLMs) pose significant individual-level privacy risks by linking fragmented multimodal data to identifiable individua

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

NeUQI: Near-Optimal Uniform Quantization Parameter Initialization for Low-Bit LLMs

DGX agent

arXiv:2505.17595v4 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve impressive performance across domains but face significant challenges when deployed on consumer-grade GPU

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm

DGX agent

arXiv:2605.31421v1 Announce Type: cross Abstract: In this paper, we show the possibility of a direct injection of algorithms into neural network architecture. We focus on a complex algorithm, that is,

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

NGDBench: Towards Neural Graph Data Management

DGX agent

arXiv:2603.05529v2 Announce Type: replace-cross Abstract: Data critical to real-world decision-making is increasingly found within organizations. Such data is heterogeneous, constantly evolving, and o

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Not All Synthetic Data Is Yours to Learn From

DGX agent

arXiv:2605.31126v1 Announce Type: cross Abstract: Can a language model improve from plain text sampled from itself, with no prompts, no teacher, no verifier, and no reward model? Yes, but only when th

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

nuReasoning: A Reasoning-Centric Dataset and Benchmark for Long-Tail Autonomous Driving

DGX agent

arXiv:2605.31572v1 Announce Type: new Abstract: Reasoning is essential for autonomous driving (AD) in long-tail scenarios, where vehicles must apply commonsense knowledge, understand spatial relations

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

OBCache: Optimal Brain KV Cache Pruning for Efficient Long-Context LLM Inference

DGX agent

arXiv:2510.07651v2 Announce Type: replace-cross Abstract: Large language models (LLMs) with extended context windows enable powerful applications but impose significant memory overhead, as caching all

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative Learning

DGX agent

arXiv:2605.30969v1 Announce Type: new Abstract: Text-based human motion editing aims to modify existing motion sequences according to natural language instructions while maintaining the consistency of

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

On-Device Generative AI for GDPR-Compliant Visual Monitoring: Natural Language Alerts from Local Object Detection

DGX agent

arXiv:2605.30544v1 Announce Type: new Abstract: Visual monitoring systems that rely on cloud-based AI inference expose raw image data to external services, creating fundamental tensions with the data-

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

DGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Pairwise Reference Alignment as a Model-Level Ordinal Observable

DGX agent

arXiv:2605.30758v1 Announce Type: new Abstract: Pairwise preference data is widely used in language-model evaluation and alignment, often for model ranking, reward modeling, or preference optimization

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Parameter-free Dynamic Regret: Time-varying Movement Costs, Delayed Feedback, and Memory

DGX agent

arXiv:2602.06902v2 Announce Type: replace Abstract: In this paper, we study dynamic regret in unconstrained online convex optimization (OCO) with movement costs. Specifically, we generalize the standa

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

PhyDrawGen: Physically Grounded Diagram Generation from Natural Language

DGX agent

arXiv:2605.30512v1 Announce Type: new Abstract: Generating physics diagrams from text requires strict adherence to physical laws. While current generative models produce visually plausible outputs, th

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Physically Viable World Models: A Case for Query-Conditioned Embodied AI

DGX agent

arXiv:2605.30542v1 Announce Type: new Abstract: World models for embodied AI must be physically viable: constructed to answer intervention queries by representing the physical structure governing acti

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Physics Enhanced Deep Surrogates for the Phonon Boltzmann Transport Equation

DGX agent

arXiv:2512.05976v3 Announce Type: replace-cross Abstract: Designing materials with controlled heat flow at the nano-scale is central to advances in microelectronics, thermoelectrics, and energy-conver

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

PInVerify: An Offline Embodied Benchmark for Active Instance Verification

DGX agent

arXiv:2605.30639v1 Announce Type: cross Abstract: Embodied agents have made strong progress in navigating to target objects, but reaching the goal vicinity does not guarantee that the agent has found

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Plain Transformers are Surprisingly Powerful Link Predictors

DGX agent

arXiv:2602.01553v2 Announce Type: replace-cross Abstract: Link prediction is a core challenge in graph machine learning, demanding models that capture rich and complex topological dependencies. While

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

PRISM: Progressive Reasoning through Iterative Slot Memory for Vision

DGX agent

arXiv:2605.30942v1 Announce Type: new Abstract: Modern vision models process images in a single feed-forward pass, which limits their ability to recover missing evidence or refine uncertain representa

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Probabilistic Precipitation Nowcasting with Rectified Flow Transformers

DGX agent

arXiv:2605.31204v1 Announce Type: new Abstract: Accurate weather forecasts are essential across various domains and are safety-critical in extreme weather conditions. Compared to simulation-based fore

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration

DGX agent

arXiv:2605.31196v1 Announce Type: cross Abstract: Safe human--robot collaboration requires more than visual description: a monitor must determine whether the robot body is safely separated, already co

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Probing the Prompt KV Cache: Where It Becomes Dispensable

DGX agent

arXiv:2605.30574v1 Announce Type: new Abstract: Prior KV cache compression schemes empirically demonstrate that the prompt cache is partially redundant during decoding, dropping or summarising entries

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

QASM-Eval: A Dataset to Train and Evaluate LLMs on OpenQASM-3 Beyond Quantum Circuits

DGX agent

arXiv:2605.30358v1 Announce Type: new Abstract: Quantum computing remains in the Noisy Intermediate-Scale Quantum (NISQ) era, where the performance is highly constrained to noise. Addressing the limit

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Quantifying the Uncertainty of Foundation Models with Singular Value Ensembles

DGX agent

arXiv:2601.22068v2 Announce Type: replace Abstract: Foundation models have become a dominant paradigm in machine learning, achieving remarkable performance across diverse tasks through large-scale pre

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Query-focused and Memory-aware Reranker for Long Context Processing

DGX agent

arXiv:2602.12192v3 Announce Type: replace Abstract: Built upon the existing analysis of retrieval heads in large language models, we propose an alternative reranking framework that trains models to es

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

QVGGT: Post-Training Quantized Visual Geometry Grounded Transformer

DGX agent

arXiv:2605.31124v1 Announce Type: new Abstract: Estimating 3D attributes directly from images has advanced rapidly with the Visual Geometry Grounded Transformer (VGGT), which predicts camera parameter

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Randomized Feasibility Methods for Constrained Optimization with Adaptive Step Sizes

DGX agent

arXiv:2601.20076v2 Announce Type: replace-cross Abstract: We consider minimizing an objective function subject to constraints defined by the intersection of lower-level sets of convex functions. We st

model-releasesarxiv-cs-lg
1 Jun 2026
← Previous
1…176177178179180…361
Next →