AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

TENP: Trapezoidal Expert Neuron Pruning For Mixture-of-Experts

DGX agent

arXiv:2606.09885v1 Announce Type: new Abstract: Mixture-of-Experts large language models (LLMs) scale efficiently through sparse activation, yet their deployment is fundamentally constrained by the la

model-releasesarxiv-cs-lg
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

The 1st PortraitCraft Challenge: A CVPR 2026 Workshop Competition on Portrait Composition Understanding and Generation

DGX agent

arXiv:2606.10894v1 Announce Type: new Abstract: This paper presents an overview of the inaugural PortraitCraft Challenge, held as one of the official competitions at CVPR 2026. The challenge focuses o

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

The hyper-scaled NLP bound for maximum-entropy remote sampling

DGX agent

arXiv:2601.20970v3 Announce Type: replace-cross Abstract: The maximum-entropy remote sampling problem (MERSP) is to select a subset of s random variables from a set of n random variables, so as to max

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

The Order Matters: Sequential Fine-Tuning of LLaMA for Coherent Automated Essay Scoring

DGX agent

arXiv:2606.10327v1 Announce Type: new Abstract: Automated Essay Scoring (AES) systems must judge interdependent discourse elements (e.g., lead, claim, evidence, conclusion), yet most approaches treat

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

The Shibboleth Effect: Auditing the Cross-Lingual Distributional Skew of Large Language Models

DGX agent

arXiv:2606.11082v1 Announce Type: new Abstract: This study investigates cross-lingual distributional skew (the Shibboleth Effect) in frontier large language models (LLMs) subjected to sustained advers

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Trainable Smooth-Rotation Transforms with Learned Channel Scales for LLM Quantization

DGX agent

arXiv:2606.09927v1 Announce Type: cross Abstract: Post-training quantization (PTQ) is one of the most practical ways to reduce the serving cost of Large Language Models (LLMs), but activation quantiza

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Training LLMs to Enforce Multi-Level Instruction Hierarchies via Gravity-Weighted Direct Preference Optimization

DGX agent

arXiv:2606.10860v1 Announce Type: cross Abstract: Production LLMs receive instructions from sources with very different levels of trust, yet attend to every token with uniform architectural privilege.

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

TRAPS: Therapeutic Response Analysis via Pathway-informed Stratification

DGX agent

arXiv:2606.09898v1 Announce Type: new Abstract: Cancer treatment planning requires decisions across multiple clinical dimensions at once. Clinicians must determine whether a patient should receive tar

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

U-TTT: Towards Generalizable PET Image Denoising via Test-Time Training

DGX agent

arXiv:2606.11032v1 Announce Type: new Abstract: Existing deep learning models for Positron Emission Tomography (PET) image denoising often suffer from severe performance degradation under distribution

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

UMI-Bench 1.0: An Open and Reproducible Real-World Benchmark for Tabletop Robotic Manipulation with UMI Data

DGX agent

arXiv:2606.10382v1 Announce Type: new Abstract: Real-robot evaluation is essential for understanding whether learned manipulation policies can operate reliably outside curated demonstrations. This nee

model-releasesarxiv-cs-ro
10 Jun 2026
Model Releases

Uncertainty-Aware Motion Planning for Autonomous Driving in Mixed Traffic Environment

DGX agent

arXiv:2606.09958v1 Announce Type: cross Abstract: In mixed-traffic environments where autonomous and human-driven vehicles may co-exist, motion planning for autonomous vehicles requires anticipating t

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

UXBench: Benchmarking User Experience in AI Assistants

DGX agent

arXiv:2606.09570v2 Announce Type: replace Abstract: As AI assistants serve millions of users daily, evaluating user experience (UX) beyond general model capability has become increasingly important. W

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions

DGX agent

arXiv:2512.11995v2 Announce Type: replace-cross Abstract: While many vision-language models (VLMs) are developed to answer well-defined, straightforward questions with highly specified targets, as in

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Validation-Stage Combinatorial Fusion Analysis for Imbalanced Credit-Card Fraud Detection

DGX agent

arXiv:2606.10393v1 Announce Type: new Abstract: Credit-card fraud detection is difficult because fraudulent transactions are rare, costly, and unevenly distributed. Strong gradient-boosted tree models

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

WebChallenger: A Reliable and Efficient Generalist Web Agent

DGX agent

arXiv:2606.10423v1 Announce Type: new Abstract: Autonomous web navigation remains challenging for LLM agents, and the strongest generalist systems rely on proprietary reasoning models whose inference

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

What Demonstration Curation Metrics Do to Your Policy

DGX agent

arXiv:2606.10229v1 Announce Type: cross Abstract: We study whether demonstration-curation metrics that detect defective training episodes also improve the downstream behavior-cloning policy that train

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

DGX agent

arXiv:2606.11045v1 Announce Type: new Abstract: Reusing a held-out benchmark adaptively should, in principle, invite overfitting. Yet benchmark-driven machine learning (ML) has produced surprisingly l

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

What makes a harness a harness: necessary and sufficient conditions for an agent harness

DGX agent

arXiv:2606.10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence. It names the layer that wraps a language

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents

DGX agent

arXiv:2606.10267v1 Announce Type: cross Abstract: Hierarchical vision-language-action (Hi-VLA) systems have emerged as a promising paradigm for complex robot manipulation, by using high-level VLM plan

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

When Design Rules Break: Benchmark Composition Determines Whether Label Informativeness Predicts GNN Aggregator Choice

DGX agent

arXiv:2606.10249v1 Announce Type: new Abstract: We examine whether graph neural network (GNN) design rules generalize across benchmark families by studying aggregator selection (sum, mean, max) on 24

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

When Do Autoregressive Sequence Models Forecast Physical Wavefields? A Controlled Study on Synthetic Seismograms

DGX agent

arXiv:2606.10868v1 Announce Type: new Abstract: Long-horizon autoregressive forecasting of oscillatory physical signals, such as seismograms, gravitational-wave strain, and similar wavefields is limit

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

When RL Fails after SFT: Rejuvenating Model Plasticity for Robust SFT-to-RL Handoff

DGX agent

arXiv:2606.09932v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) has become a standard pipeline for Large Language Model (LLM) post-training. SFT

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Who Brought Easter Eggs to Eid? Auditing Cultural Translation of Math Word Problems Across Diverse Languages and Regions

DGX agent

arXiv:2606.11009v1 Announce Type: new Abstract: Large language models are increasingly used to adapt math word problems for personalized learning at scale, but it remains an open question whether thos

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

WHU-Infra3D: A Full-stack Multi-modal Dataset and Benchmark for 3D Roadside Infrastructure Inventory

DGX agent

arXiv:2606.09882v1 Announce Type: new Abstract: The paradigm of digital twin cities is shifting from coarse visual mapping toward more precise and actionable digitization of urban assets. However, exi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields

DGX agent

arXiv:2606.11042v1 Announce Type: new Abstract: Recent years have witnessed the rapid evolution of AI agents toward handling increasingly complex, real-world tasks. However, existing benchmarks rarely

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

WorldOlympiad: Can Your World Model Survive a Triathlon?

DGX agent

arXiv:2606.11129v1 Announce Type: new Abstract: We introduce WorldOlympiad, a benchmark for diagnosing video-based world models across physical faithfulness, geometric consistency, and interaction fid

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

XtrAIn: Training-Guided Occlusion for Feature Attribution

DGX agent

arXiv:2606.10877v1 Announce Type: cross Abstract: Occlusion-based attribution methods provide an intuitive way to estimate feature importance by perturbing input features and measuring the resulting c

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

A Baseline Study and Benchmark for Few-Shot Open-Set Action Recognition with Feature Residual Discrimination

DGX agent

arXiv:2603.04125v2 Announce Type: replace Abstract: Few-Shot Action Recognition (FS-AR) has shown promising results but is often limited by a closed-set assumption that fails in real-world open-set sc

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

A Comparative Study of Student Perspectives on Technical Writing Feedback Quality: Evaluating LLMs, SLMs, and Humans in Computer Science Topics

DGX agent

arXiv:2601.11541v2 Announce Type: replace-cross Abstract: To address the scalability of feedback in computer science while mitigating the privacy and cost limitations of commercial Large Language Mode

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A Comparison of SSL-Based Feature Extractors and Back-End Classifiers for Spoofing Detection: A Multi-Corpus Training and Cross-Linguistic Analysis

DGX agent

arXiv:2606.08669v1 Announce Type: cross Abstract: Voice biometric systems face growing threats from spoofing attacks, yet the evaluation of detection models remains inconsistent across datasets. To in

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

A Dataset for Dynamic Human Preferences for Vision Language Models

DGX agent

arXiv:2606.07653v1 Announce Type: cross Abstract: Given the increased adoption of Vision Language Models (VLMs) in human-interactive settings, it is important that we evaluate how well these models ca

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A Framework for Evaluating and Benchmarking Concept Drift Detection Methods

DGX agent

arXiv:2606.07789v1 Announce Type: new Abstract: Data stream mining is fundamentally challenged by concept drift, where distributional changes can degrade model performance. Despite the proliferation o

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology

DGX agent

arXiv:2606.08093v1 Announce Type: new Abstract: Pathology is the cornerstone of modern medicine, where accurate decision-making relies heavily on evidence-based practices. While artificial intelligenc

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A retrieval conditioned rebinding circuit for dynamic entity tracking in large language models

DGX agent

arXiv:2606.08644v1 Announce Type: cross Abstract: To interpret context correctly and retrieve relevant information, large language models must bind entities to their attributes and update these bindin

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A Survey of Heterogeneous Graph Neural Networks for Cybersecurity Anomaly Detection

DGX agent

arXiv:2510.26307v3 Announce Type: replace-cross Abstract: Anomaly detection is a critical task in cybersecurity, where identifying insider threats, access violations, and coordinated attacks is essent

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

A Unifying Framework for Concept-Based Representational Similarity

DGX agent

arXiv:2606.09653v1 Announce Type: new Abstract: Learned representations across models and modalities often exhibit striking structural similarities, suggesting shared underlying concept decompositions

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

ABLE: Representing and Mapping LLMs via Attribution-Based Large-model Embedding

DGX agent

arXiv:2606.07524v1 Announce Type: cross Abstract: The explosive growth of large language models (LLMs) has created a heterogeneous and poorly documented ecosystem, making systematic model comparison i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Activation Steering Induces Emergent Misalignment: A More Comprehensive Evaluation

DGX agent

arXiv:2606.08682v1 Announce Type: cross Abstract: Activation steering has emerged as a popular inference-time technique for modulating the behavior of large language models (LLMs). By constructing a s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning

DGX agent

arXiv:2505.21457v2 Announce Type: replace-cross Abstract: Active vision, also known as active perception, refers to actively selecting where and how to look in order to gather task-relevant informatio

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Adaptive directional gradients for parameterised quantum circuits

DGX agent

arXiv:2606.09734v1 Announce Type: cross Abstract: Training parameterised quantum circuits (PQCs) on quantum hardware is bottlenecked by the measurement cost of gradient estimation, which under the par

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

AgentCompile: An LLM-Guided Compiler for Direct CUDA Inference

DGX agent

arXiv:2606.07665v1 Announce Type: cross Abstract: Transformer inference increasingly depends on specialized compiler and runtime support, but real model graphs still require semantic decisions about w

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

AgroOmni: A Large-Scale Multi-view Agricultural Dataset for Cross-Scale Multimodal Reasoning

DGX agent

arXiv:2603.14342v2 Announce Type: replace-cross Abstract: Modern agricultural data is sourced from diverse platforms and spans multiple spatial scales, ranging from ground-level close-up photography t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

AI Scientists Are Only as Good as Their Evidence: A Stratified Ablation of Proprietary Data and Reasoning Skills in Drug-Asset Valuation

DGX agent

arXiv:2606.09556v1 Announce Type: new Abstract: AI Scientist agents are often evaluated as if capability were mainly a function of model quality, prompting, or reasoning scaffolds. We test a different

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Alcmean's: Unsupervised community detection using local Laplacian, automatic detection of the number of centers

DGX agent

arXiv:2606.09100v1 Announce Type: cross Abstract: Community detection is a fundamental problem in the analysis of complex networks. It has applications across social, biological, and financial domains

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

AliyunConsoleAgent: Training Web Agents in Real-World Cloud Environments via Distillation and Reinforcement Learning

DGX agent

arXiv:2606.09447v1 Announce Type: new Abstract: We present AliyunConsoleAgent, a web agent framework for automated documentation verification in real-world cloud consoles. Major cloud platforms encomp

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

AlphaOPT: Formulating Optimization Programs with Self-Improving LLM Experience Library

DGX agent

arXiv:2510.18428v4 Announce Type: replace Abstract: Optimization modeling underlies critical decision-making across industries, yet remains difficult to automate: natural-language problem descriptions

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

AMN: An Adaptive Multi-Scale Fusion Network with Boundary and Uncertainty Modeling for Nuclei Segmentation

DGX agent

arXiv:2606.07633v1 Announce Type: cross Abstract: Accurate classification of nuclei subtypes in histopathology images is critical for downstream tasks including tumor grading, immune infiltrate quanti

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…139140141142143…361
Next →