AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Beyond Heuristics: Learnable Density Control for 3D Gaussian Splatting

DGX agent

arXiv:2605.00408v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has demonstrated impressive real-time rendering performance, its efficacy remains constrained by a reliance on heuris

model-releasesarxiv-cs-cv
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration

DGX agent

arXiv:2605.00310v1 Announce Type: new Abstract: Super-resolution (SR) techniques have made major advances in reconstructing high-resolution images from low-resolution inputs. The increased resolution

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Borrowed Geometry: Computational Reuse of Frozen Text-Pretrained Transformer Weights Across Modalities

DGX agent

arXiv:2605.00333v1 Announce Type: cross Abstract: Frozen Gemma 4 31B weights pretrained exclusively on text tokens, unmodified, transfer across modality boundaries through a thin trainable interface.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Bring Your Own Prompts: Use-Case-Specific Bias and Fairness Evaluation for LLMs

DGX agent

arXiv:2407.10853v5 Announce Type: replace Abstract: Bias and fairness risks in Large Language Models (LLMs) vary substantially across deployment contexts, yet existing approaches lack systematic guida

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Can Coding Agents Reproduce Findings in Computational Materials Science?

DGX agent

arXiv:2605.00803v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous coding agents and have achieved remarkably strong performance on software engineering be

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Caracal: Causal Architecture via Spectral Mixing

DGX agent

arXiv:2605.00292v1 Announce Type: new Abstract: The scalability of Large Language Models to long sequences is hindered by the quadratic cost of attention and the limitations of positional encodings. T

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection

DGX agent

arXiv:2605.00630v1 Announce Type: new Abstract: The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

CompleteRXN: Toward Completing Open Chemical Reaction Databases

DGX agent

arXiv:2605.00222v1 Announce Type: new Abstract: Chemical reaction datasets such as USPTO suffer from substantial incompleteness, frequently missing byproducts, co-reactants, and stoichiometric coeffic

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Concolic Testing on Individual Fairness of Neural Network Models

DGX agent

arXiv:2509.06864v2 Announce Type: replace Abstract: This paper introduces PyFair, a formal framework for evaluating and verifying individual fairness of Deep Neural Networks (DNNs). By adapting the co

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

ControBench: An Interaction-Aware Benchmark for Controversial Discourse Analysis on Social Networks

DGX agent

arXiv:2605.00513v1 Announce Type: new Abstract: Understanding how people argue across ideological divides online is important for studying political polarization, misinformation, and content moderatio

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

CURE-OOD: Benchmarking Out-of-Distribution Detection for Survival Prediction

DGX agent

arXiv:2605.00350v1 Announce Type: new Abstract: ``How long can I live and remain free of cancer?'' is often the first question a patient asks after receiving a cancer diagnosis and treatment. Accurate

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Deepfakes: we need to re-think the concept of 'real' images

DGX agent

arXiv:2509.21864v2 Announce Type: replace Abstract: The wide availability and low usability barrier of modern image generation models has triggered the reasonable fear of criminal misconduct and negat

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Differentiable Autoencoding Neural Operator for Interpretable and Integrable Latent Space Modeling

DGX agent

arXiv:2510.00233v2 Announce Type: replace Abstract: Scientific machine learning has enabled the extraction of physical insights and data-driven modeling of high-dimensional spatiotemporal data, yet ac

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Do Open-Loop Metrics Predict Closed-Loop Driving? A Cross-Benchmark Correlation Study of NAVSIM and Bench2Drive

DGX agent

arXiv:2605.00066v1 Announce Type: new Abstract: Open-loop evaluation offers fast, reproducible assessment of autonomous driving planners, but its ability to predict real closed-loop driving performanc

model-releasesarxiv-cs-ro
4 May 2026
Model Releases

Driving with A Thousand Faces: A Benchmark for Closed-Loop Personalized End-to-End Autonomous Driving

DGX agent

arXiv:2602.18757v2 Announce Type: replace Abstract: Human driving behavior is inherently diverse, yet most end-to-end autonomous driving (E2E-AD) systems learn a single average driving style, neglecti

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Dynamics-Encoded Deep Learning for Robust System Identification and Parameter Estimation

DGX agent

arXiv:2410.04299v2 Announce Type: replace Abstract: Incorporating a priori physics knowledge into machine learning leads to more robust and interpretable algorithms. In this work, we combine deep lear

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Efficient Spatio-Temporal Vegetation Pixel Classification with Vision Transformers

DGX agent

arXiv:2605.00296v1 Announce Type: new Abstract: Plant phenology-the study of recurrent life cycle events-is essential for understanding ecosystem dynamics and their responses to climate change impacts

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game

DGX agent

arXiv:2605.00677v1 Announce Type: new Abstract: While Large Language Models have achieved notable success on formal mathematics benchmarks such as MiniF2F, it remains unclear whether these results ste

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Event-based Civil Infrastructure Visual Defect Detection: ev-CIVIL Dataset and Benchmark

DGX agent

arXiv:2504.05679v2 Announce Type: replace Abstract: Small unmanned aerial vehicle (UAV)-based visual inspections are a more efficient alternative to manual methods for examining civil structural defec

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation

DGX agent

arXiv:2507.14201v3 Announce Type: replace-cross Abstract: We present ExCyTIn-Bench, the first benchmark to Evaluate an LLM agent X on the task of Cyber Threat Investigation through security questions

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Exploring the System 1 Thinking Capability of Large Reasoning Models

DGX agent

arXiv:2504.10368v4 Announce Type: replace Abstract: This paper explores the system 1 thinking capability of Large Reasoning Models (LRMs), the intuitive ability to respond efficiently with minimal tok

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Faithful Extreme Image Rescaling with Learnable Reversible Transformation and Semantic Priors

DGX agent

arXiv:2605.00605v1 Announce Type: new Abstract: Most recent extreme rescaling methods struggle to preserve semantically consistent structures and produce realistic details, due to the severely ill-pos

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

FedACT: Concurrent Federated Intelligence across Heterogeneous Data Sources

DGX agent

arXiv:2605.00011v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative intelligence across decentralized data source devices in a privacy-preserving way. While substantial resea

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios

DGX agent

arXiv:2605.00706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in financial scenarios. However, they may produce harmful outputs, including facilitating illegal

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

FollowTable: A Benchmark for Instruction-Following Table Retrieval

DGX agent

arXiv:2605.00400v1 Announce Type: cross Abstract: Table Retrieval (TR) has traditionally been formulated as an ad-hoc retrieval problem, where relevance is primarily determined by topical semantic sim

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Foresight Arena: An On-Chain Benchmark for Evaluating AI Forecasting Agents

DGX agent

arXiv:2605.00420v1 Announce Type: cross Abstract: Evaluating the true forecasting ability of AI agents requires environments resistant to overfitting, free from centralized trust, and grounded in ince

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

From Backward Spreading to Forward Replay: Revisiting Target Construction in LLM Parameter Editing

DGX agent

arXiv:2605.00358v1 Announce Type: new Abstract: LLM parameter editing methods commonly rely on computing an ideal target hidden-state at a target layer (referred as anchor point) and distributing the

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

From Images2Mesh: A 3D Surface Reconstruction Pipeline for Non-Cooperative Space Objects

DGX agent

arXiv:2605.00147v1 Announce Type: new Abstract: On-orbit inspection imagery is crucial as it enables characterization of non-cooperative resident space objects, providing the geometry and structural c

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

From Prediction to Practice: A Task-Aware Evaluation Framework for Blood Glucose Forecasting

DGX agent

arXiv:2605.00645v1 Announce Type: new Abstract: Clinical time-series forecasting is increasingly studied for decision support, yet standard aggregate metrics can obscure whether a model is actually us

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking

DGX agent

arXiv:2605.00281v1 Announce Type: new Abstract: We study high-probability (HP) convergence guarantees in decentralized stochastic optimization, where multiple agents collaborate to jointly train a mod

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

How Frontier LLMs Adapt to Neurodivergence Context: A Measurement Framework for Surface vs. Structural Change in System-Prompted Responses

DGX agent

arXiv:2605.00113v1 Announce Type: new Abstract: We examine if frontier chat-based large language models (LLMs) adjust their outputs based on neurodivergence (ND) context in system prompts and describe

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks

DGX agent

arXiv:2507.01955v3 Announce Type: replace Abstract: Multimodal foundation models (MFMs), such as GPT-4o, have recently made remarkable progress. However, their detailed visual understanding beyond que

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Hypergraph and Latent ODE Learning for Multimodal Root Cause Localization in Microservices

DGX agent

arXiv:2605.00351v1 Announce Type: new Abstract: Root cause localization in cloud native microservice systems requires modeling complex service dependencies, irregular temporal dynamics, and heterogene

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information

DGX agent

arXiv:2508.07630v2 Announce Type: replace Abstract: We introduce InterChart, a diagnostic benchmark that evaluates how well vision-language models (VLMs) reason across multiple related charts, a task

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Introducing WARM-VR: Benchmark Dataset for Multimodal Wearable Affect Recognition in Virtual Reality

DGX agent

arXiv:2605.00184v1 Announce Type: new Abstract: With the growing integration of human-computer interaction into everyday life, advances in machine learning have enabled systems to better perceive and

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Jailbreaking Vision-Language Models Through the Visual Modality

DGX agent

arXiv:2605.00583v1 Announce Type: new Abstract: The visual modality of vision-language models (VLMs) is an underexplored attack surface for bypassing safety alignment. We introduce four jailbreak atta

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Jailbroken Frontier Models Retain Their Capabilities

DGX agent

arXiv:2605.00267v1 Announce Type: new Abstract: As language model safeguards become more robust, attackers are pushed toward developing increasingly complex jailbreaks. Prior work has found that this

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Knowing When to Defer: Selective Prediction for Responsible Knowledge Tracing

DGX agent

arXiv:2509.21514v3 Announce Type: replace-cross Abstract: Research on Knowledge Tracing (KT) models traditionally focuses on improving predictive accuracy. However, responsible real-world deployment r

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Learning from Compressed CT: Feature Attention Style Transfer and Structured Factorized Projections for Resource-Efficient Medical Image Analysis

DGX agent

arXiv:2605.00448v1 Announce Type: new Abstract: The deployment of artificial intelligence in medical imaging is hindered by high computational complexity and resource-intensive processing of volumetri

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Learning from Supervision with Semantic and Episodic Memory: A Reflective Approach to Agent Adaptation

DGX agent

arXiv:2510.19897v2 Announce Type: replace Abstract: We investigate how agents built on pretrained large language models (LLMs) can learn target classification functions from labeled examples without p

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Learning from the Unseen: Generative Data Augmentation for Geometric-Semantic Accident Anticipation

DGX agent

arXiv:2605.00051v1 Announce Type: new Abstract: Anticipating traffic accidents is a critical yet unresolved problem for autonomous driving, hindered by the inherent complexity of modeling interactions

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Learning Locally, Revising Globally: Global Reviser for Federated Learning with Noisy Labels

DGX agent

arXiv:2412.00452v2 Announce Type: replace-cross Abstract: Conventioanl federated learning (FL) heavily depends on high-quality labels, which are often impractical in the real world, leading to the fed

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context

DGX agent

arXiv:2505.22003v2 Announce Type: replace Abstract: In India, access to legal assistance for the general public has been observed to have a critical gap, as many citizens are not able to take full adv

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

LLM-Oriented Information Retrieval: A Denoising-First Perspective

DGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

M-CaStLe: Uncovering Local Causal Structures in Multivariate Space-Time Gridded Data

DGX agent

arXiv:2605.00398v1 Announce Type: new Abstract: Causal graph discovery for space-time systems is challenging in high-dimensional gridded data, which often has many more grid cells than temporal observ

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Make Your LVLM KV Cache More Lightweight

DGX agent

arXiv:2605.00789v1 Announce Type: new Abstract: Key-Value (KV) cache has become a de facto component of modern Large Vision-Language Models (LVLMs) for inference. While it enhances decoding efficiency

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

DGX agent

arXiv:2510.17281v5 Announce Type: replace Abstract: Scaling up data, parameters, and test-time computation has been the mainstream methods to improve LLM systems (LLMsys), but their upper bounds are a

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Minimizing Human Intervention in Online Classification

DGX agent

arXiv:2510.23557v2 Announce Type: replace-cross Abstract: Training or fine-tuning large language model (LLM)-based systems often requires costly human feedback, yet there is limited understanding of h

model-releasesarxiv-cs-lg
4 May 2026
← Previous
1…283284285286287…361
Next →