AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Parameter-Free Encoders Remain Viable for RDB Foundation Models

DGX agent

arXiv:2607.05476v1 Announce Type: new Abstract: Given a relational database (RDB) storing heterogeneous tabular information, how can we predict missing (or future) values in some target column of inte

model-releasesarxiv-cs-lg
8 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Partial Symmetry Detection for 3D Geometry using Contrastive Learning with Geodesic Point Cloud Patches

DGX agent

arXiv:2312.08230v2 Announce Type: replace Abstract: Detecting partial extrinsic symmetry in 3D geometry is a fundamental yet persistent challenge in computer vision and graphics, critical for tasks ra

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

PatchOptic for Shared-State LLM Workflows with Projected Views and Verified Structured Updates

DGX agent

arXiv:2607.05483v1 Announce Type: cross Abstract: Agentic workflows often operate over shared, structured state. Because LLM context windows are limited, each model invocation is typically shown only

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation

DGX agent

arXiv:2607.05915v1 Announce Type: new Abstract: PCB routing is the task of connecting the nets of a board with copper traces under strict design rules, yet learning-based methods still lag behind rule

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PIPBench: A Profile-Inclusive Framework for Personalized Image Generation Evaluation

DGX agent

arXiv:2607.06440v1 Announce Type: new Abstract: Recent text-to-image models such as DALLE-3 excel at following diverse prompts yet remain blind to individual aesthetic preferences. We study personaliz

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Pluralis v0.1: Towards a Multicultural, Multimodal, Multilingual Benchmark for AI Risk and Reliability

DGX agent

arXiv:2607.06196v1 Announce Type: new Abstract: Current AI safety evaluation and benchmarking frameworks predominantly rely on Western-centric culture-agnostic defaults that mask critical regional law

model-releasesarxiv-cs-cl
8 Jul 2026
Model Releases

PluraMath: Extending Mathematical Reasoning Evaluation Beyond High-Resource Languages

DGX agent

arXiv:2607.05992v1 Announce Type: cross Abstract: Mathematical reasoning has become a central task for evaluating and tuning reasoning Large Language Models (LLMs), yet existing benchmarks remain heav

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails

DGX agent

arXiv:2607.05910v1 Announce Type: cross Abstract: Image guardrails are typically trained and evaluated under a fixed safety policy, implicitly treating safety as an intrinsic property of an image. Rea

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents

DGX agent

arXiv:2607.06008v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong performance in long-horizon tasks that require planning, tool use, and interaction with external env

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Population-Level Profiling of DSM-5 Depressive Symptoms Among Self-Reported ADHD and ASD Users on Twitter: An Exploratory Study Using Advanced NLP and Statistical Analysis

DGX agent

arXiv:2607.05626v1 Announce Type: new Abstract: Background: Depression frequently co-occurs with ADHD and autism spectrum disorder (ASD), but population-level differences in symptom expression between

model-releasesarxiv-cs-cl
8 Jul 2026
Model Releases

Privilege and confidentiality in generative AI workflows

DGX agent

arXiv:2607.05479v1 Announce Type: cross Abstract: Generative AI (GenAI) systems store and process client data in three distinct ways: in the model's parameters through training and memorisation, in th

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Prompt-Adapter Context Routing for Parameter-Efficient Multi-Shot Long Video Extrapolation

DGX agent

arXiv:2607.06481v1 Announce Type: cross Abstract: We present PACR-Video, a parameter-efficient framework for multi-shot long video extrapolation that preserves recurring entities, scene structure, vis

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Quantifying Frontier LLM Capabilities for Container Sandbox Escape

DGX agent

arXiv:2603.02277v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly act as autonomous agents, using tools to execute code, read and write files, and access networks, cr

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

RFHNet: Relational and Frequency-Aware Hashing Network for Large-Scale Fine-Grained Food Image Retrieval

DGX agent

arXiv:2607.06148v1 Announce Type: new Abstract: Fine-grained food image retrieval is a key task in computational gastronomy, with applications in food traceability, dietary monitoring, and smart cater

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

RPAM: A Principled Metric for Evaluating Associations in Language Models with High Predictive Validity in Downstream Outputs

DGX agent

arXiv:2607.05679v1 Announce Type: cross Abstract: Language models (LMs) exhibit problematic biases, such as stereotypes. Effectively analyzing and mitigating such biases requires accurate and generali

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications

DGX agent

arXiv:2607.06411v1 Announce Type: cross Abstract: Developers increasingly delegate real maintenance work to product-grade coding agents, and many state tasks in their native language, in the style of

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

SafeImpute: Reliable Clinical Data Imputation via Conformal Selection

DGX agent

arXiv:2607.05613v1 Announce Type: new Abstract: Clinical care often relies on key laboratory indicators, yet real-world patient visits are sparse and tests are ordered irregularly, leading to pervasiv

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

SAGE: Spatial-visual Adaptive Graph Exploration for Efficient Visual Place Recognition

DGX agent

arXiv:2509.25723v4 Announce Type: replace Abstract: Visual Place Recognition (VPR) requires robust retrieval of geotagged images despite large appearance, viewpoint, and environmental variation. Prior

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Scientific Code Search at Scale: A Multi-Domain Dataset and Benchmark

DGX agent

arXiv:2607.05443v1 Announce Type: cross Abstract: Scientists increasingly rely on open-source tools to support their research workflows, yet discovering relevant software among over 600 million GitHub

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Self-Review Reinforcement Learning (SRRL) with Cross-Episode Memory and Policy Distillation

DGX agent

arXiv:2607.05541v1 Announce Type: cross Abstract: Reinforcement Learning is commonly used to train large language models using environmental feedback. In applied settings, the environment usually prov

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Self-Routing: Parameter-Free Expert Routing from Hidden States

DGX agent

arXiv:2604.00421v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) layers increase model capacity by activating only a small subset of experts per token, and typically rely on a learned rout

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

SEVRA-BENCH: Social Engineering of Vulnerabilities in Review Agents

DGX agent

arXiv:2606.13757v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed in automated code-review systems, where their approvals can determine which code is mer

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Social 3D Scene Graphs: Modeling Human Actions and Relations for Interactive Service Robots

DGX agent

arXiv:2509.24966v2 Announce Type: replace Abstract: Understanding how people interact with their surroundings and each other is essential for enabling robots to act in socially compliant and context-a

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

SpanUQ: Span-Level Uncertainty Quantification for Large Language Model Generation

DGX agent

arXiv:2607.05721v1 Announce Type: new Abstract: Uncertainty estimation is essential not only for the trustworthy deployment of large language models (LLMs) but also as a foundation for self-refinement

model-releasesarxiv-cs-cl
8 Jul 2026
Model Releases

Sparse but Wrong: Incorrect L0 Leads to Incorrect Features in Sparse Autoencoders

DGX agent

arXiv:2508.16560v4 Announce Type: replace-cross Abstract: Sparse Autoencoders (SAEs) extract features from LLM internal activations, meant to correspond to interpretable concepts. A core SAE training

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows

DGX agent

arXiv:2607.06229v1 Announce Type: cross Abstract: Major cloud data platforms now expose large language model capabilities as native SQL functions, enabling analysts to perform classification, filterin

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Statistically Meaningful Geometry and Gauge Symmetry Breaking: A Geometric Foundation for Scientific Discovery and Intelligence Emergence

DGX agent

arXiv:2607.05436v1 Announce Type: new Abstract: The rapid scaling of over-parameterized machine learning architectures, particularly LLMs, raises a profound crisis: do these systems exhibit genuine in

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

StepShield: When, Not Whether to Intervene on Rogue Agents

DGX agent

arXiv:2601.22136v2 Announce Type: replace-cross Abstract: Agent safety benchmarks measure whether a monitor detects harm, not when. Yet timing is the difference between intervention and autopsy. We in

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Structured Data Extraction from Real Estate Documents using Clustering, Classification, and Large Language Models

DGX agent

arXiv:2607.06012v1 Announce Type: new Abstract: Real estate property listings expose structured metadata through the API. Still, the richest property-level information (i.e., legal status, structural

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Supervised Reward Inference

DGX agent

arXiv:2502.18447v2 Announce Type: replace Abstract: Existing approaches to reward inference typically assume that humans provide demonstrations according to specific behavior models. However, humans o

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

DGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

The Granularity Paradox: How Temporal Disaggregation Inflates In-Sample Fit and Compounds Out-of-Sample Error

DGX agent

arXiv:2607.05450v1 Announce Type: cross Abstract: This paper explores the 'Granularity Paradox' in time-series forecasting, wherein finer temporal disaggregation (e.g., Monthly to Weekly/Daily) improv

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

The relationship between reasoning and performance in large language models--o3 (mini) thinks harder, not longer

DGX agent

arXiv:2502.15631v2 Announce Type: replace-cross Abstract: Large language models have demonstrated remarkable progress in mathematical reasoning, leveraging chain-of-thought and reinforcement learning.

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

The yes-no bias of large language models reflects answer order and wording, not shifts in moral judgment

DGX agent

arXiv:2607.05552v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly issue judgments read as binary verdicts, and a growing literature reports such judgments shifting under logi

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Think Before You Grid-Search: Floor-First Triage for LLM Serving

DGX agent

arXiv:2607.05876v1 Announce Type: cross Abstract: LLM serving optimization typically benchmarks many configurations and reaches for heavy profilers when latency targets are missed. We argue for the re

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

ThorArena: Benchmarking Humanoid Physical Interaction with Human Motion-Force Demonstrations

DGX agent

arXiv:2607.06052v1 Announce Type: new Abstract: Humanoid robots are increasingly expected to perform contact-rich tasks that require not only accurate whole-body motion but also robust physical intera

model-releasesarxiv-cs-ro
8 Jul 2026
Model Releases

Transformers converge to invariant algorithmic cores

DGX agent

arXiv:2602.22600v2 Announce Type: replace-cross Abstract: Training selects for behavior, not circuitry: many weight configurations can implement the same function. Studying any single trained neural n

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

UI2App: Benchmarking Visual Interaction Inference in Executable Web Application Generation

DGX agent

arXiv:2607.06306v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex pro

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

UniLM-Nav: A Unified Framework for Zero-Shot Last-Mile Navigation

DGX agent

arXiv:2607.06537v1 Announce Type: new Abstract: Mobile manipulation requires a robot to navigate to a target object or receptacle and then perform intended manipulation. However, reaching the vicinity

model-releasesarxiv-cs-ro
8 Jul 2026
Model Releases

VendorBench-100: A Unified Cross-Paradigm Benchmark for Deepfake Image Detection

DGX agent

arXiv:2607.06254v1 Announce Type: cross Abstract: Deepfake image detection is currently served by three fundamentally different paradigms: commercial APIs, zero-shot vision-language models (LLMs), and

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Verification of Dynamic Holographic Behavior in Identity Documents

DGX agent

arXiv:2607.06466v1 Announce Type: new Abstract: This paper addresses the remote verification of the authenticity of Optically Variable Devices (commonly known as holograms) on identity documents. Typi

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

VisCoP: Visual Probing for Video Domain Adaptation of Vision Language Models

DGX agent

arXiv:2510.13808v2 Announce Type: replace Abstract: Large Vision Language Models (VLMs) excel at general visual reasoning but experience significant performance degradation when deployed in novel doma

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

WebRetriever: A Large-Scale Comprehensive Benchmark for Efficient Web Agent Evaluation

DGX agent

arXiv:2607.06118v1 Announce Type: new Abstract: As web agents increasingly demonstrate capabilities in automated task execution, the development of robust evaluation frameworks for assessing their nav

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

When Should LLMs Search? Counterfactual Supervision for Search Routing

DGX agent

arXiv:2607.05752v1 Announce Type: cross Abstract: Search-augmented language models can use external evidence to compensate for limitations in parametric knowledge, but search is not uniformly benefici

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Why does Deep Learning Improve Visual SLAM?

DGX agent

arXiv:2607.06023v1 Announce Type: new Abstract: Visual SLAM is a well-established technology utilized in a wide range of real-world applications. However, its performance still degrades under challeng

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

x-Prediction Is All You Need:Training-Free Accelerated Generation via Endpoint Decodability

DGX agent

arXiv:2607.06114v1 Announce Type: cross Abstract: Diffusion and flow matching models generate high-quality samples, but their ODE samplers often need tens to hundreds of neural function evaluations (N

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

XRFormer: Multiscale Tokenization for XRF Representation Learning

DGX agent

arXiv:2607.06424v1 Announce Type: new Abstract: X-ray fluorescence (XRF) spectroscopy is a key modality for material analysis in cultural heritage. However, automated learning from XRF spectra remains

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models

DGX agent

arXiv:2502.09696v3 Announce Type: replace Abstract: Large Multimodal Models (LMMs) exhibit shortfalls when interpreting images and, by some measures, have poorer spatial cognition than young children

model-releasesarxiv-cs-cv
8 Jul 2026
← Previous
1…8283848586…361
Next →