AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

APEX4: Efficient Pure W4A4 LLM Inference via Intra-SM Compute Rebalancing

DGX agent

arXiv:2606.08761v1 Announce Type: cross Abstract: W4A4 quantization promises full utilization of INT4 Tensor Cores, yet group dequantization overhead on CUDA Cores has driven existing systems to mixed

model-releasesarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Are Reasoning Vision-Language Models Robust to Semantic Visual Distractions?

DGX agent

arXiv:2606.08894v1 Announce Type: new Abstract: Reasoning Vision-Language Models (VLMs) achieve strong performance on complex multimodal tasks, but reliable real-world application requires handling vi

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

ArtiFact: A Large-Scale Multi-Modal Cultural Heritage Dataset

DGX agent

arXiv:2606.09648v1 Announce Type: cross Abstract: Multi-modal data management has emerged as a central research topic in the database community, spanning data integration, semantic query processing, a

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery

DGX agent

arXiv:2606.08728v1 Announce Type: new Abstract: Mathematical reasoning has long served as a stringent test of machine intelligence; over the past decade, it has moved from a niche problem within NLP t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ATN3D: Density-Aware LiDAR-Radar Early 3D Object Detection Under Extreme Sparsity

DGX agent

arXiv:2606.09634v1 Announce Type: cross Abstract: 3D object detection is the backbone of perception for automated vehicles (AV) and broader intelligent transportation systems applications. Long-range

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Auditable Graph-Guided Root Cause Analysis for Kubernetes Incidents

DGX agent

arXiv:2606.08590v1 Announce Type: cross Abstract: Kubernetes incidents are diagnosed reliably only when a root-cause system's reported gains come from incident evidence rather than scenario-specific s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Auditing Proprietary Alignment in Large Language Models: A Comparative Framework Without a Ground-Truth Standard

DGX agent

arXiv:2606.08381v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly released and deployed through opaque development and deployment pipelines, enabling model providers to i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Automatic Extraction of Structured Information from Brain MRI Reports Using an Open-Weight Large Language Model

DGX agent

arXiv:2606.07721v1 Announce Type: new Abstract: Objectives: Automatic data extraction from free-text radiology reports enables large-scale research, but few studies assessed the performance of large l

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

AutoMegaKernel: A Statically-Checked Agent Harness for Self-Retargeting Megakernel Synthesis

DGX agent

arXiv:2606.09682v1 Announce Type: new Abstract: AutoMegaKernel (AMK) compiles a HuggingFace Llama-family model into a single persistent cooperative CUDA kernel that runs the whole forward pass in one

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs

DGX agent

arXiv:2606.07643v1 Announce Type: cross Abstract: Recent advances in Omni-Multimodal Large Language Models (Omni-MLLMs) have enabled strong integration of vision, audio, and language. However, their a

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Bayesian Optimization of a Multi-Product Chemical Reactor Using Composite Models and Partial Physics Knowledge

DGX agent

arXiv:2606.08611v1 Announce Type: cross Abstract: We study data-driven real-time economic optimization of a multi-product chemical reactor when no reliable first-principles model is available beyond a

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Bayesian Selective Latent Inference for Wastewater-First Influenza Monitoring

DGX agent

arXiv:2606.09433v1 Announce Type: new Abstract: Wastewater influenza surveillance can reveal community circulation before clinical reporting, but wastewater alone is not a fully identifiable proxy for

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Benchmark Datasets for Lead-Lag Forecasting on Social Platforms

DGX agent

arXiv:2511.03877v2 Announce Type: replace Abstract: Social and collaborative platforms emit multivariate time-series traces in which early interactions -- such as views, likes, or downloads -- are fol

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Benchmarking Empirical Privacy Protection for Adaptations of Large Language Models

DGX agent

arXiv:2606.09401v1 Announce Type: new Abstract: Recent work has applied differential privacy (DP) to adapt large language models (LLMs) for sensitive applications, offering theoretical guarantees. How

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Benchmarking Open-Ended Multi-Agent Coordination in Language Agents

DGX agent

arXiv:2606.08340v1 Announce Type: new Abstract: As language models are increasingly deployed as autonomous agents, they must coordinate with others over long horizons in open-ended interactive tasks.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Benchmarking Vision-Language-Action Models on SO-101: Failure and Recovery Analysis

DGX agent

arXiv:2606.08881v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong generalization in robotic manipulation, yet existing evaluations are primarily conducted

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Consistency: Preserving Temporal Structure in Zero-Shot Video Editing

DGX agent

arXiv:2606.08780v1 Announce Type: new Abstract: Existing zero-shot video editing methods rely on pre-trained diffusion models, successfully achieving spatial control and basic temporal consistency but

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Beyond English benchmarks: clinical llm evaluation in Brazilian Portuguese

DGX agent

arXiv:2606.07853v1 Announce Type: cross Abstract: Large Language Models are transforming the support for clinical decision and their application in real scenarios. Yet, most benchmarks are conducted i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

DGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Pass Rate: A Multilingual, Execution-Grounded Evaluation of Open Code LLMs

DGX agent

arXiv:2606.08840v1 Announce Type: new Abstract: Code generation models are typically compared using compact execution benchmarks and aggregate pass rates, but such summaries obscure how performance va

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Pass/Fail: Using Process Mining to Understand How LLMs Resist (and Fail) Red Team Attacks

DGX agent

arXiv:2606.07833v1 Announce Type: cross Abstract: Standard AI red teaming evaluations reduce adversarial campaigns to a single binary outcome, attack success rate (ASR), not taking into account the se

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Bidirectional Semantic Complementary Tool Retrieval for Remote Sensing Agents

DGX agent

arXiv:2606.07538v1 Announce Type: cross Abstract: Large language model (LLM)-based agents provide a novel paradigm for the automated processing of remote sensing(RS) data. Their success in complex RS

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Bidirectional Small-Granularity Search between Code and Text

DGX agent

arXiv:2606.07519v1 Announce Type: cross Abstract: We introduce the novel task of bidirectional small-granularity search between code and text, where the queries are small snippets of text or code and

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension

DGX agent

arXiv:2606.08674v1 Announce Type: cross Abstract: Existing video generation frameworks treat sequence duration as an externally prescribed parameter -- fixed frame counts or text prompts -- producing

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

BLUE: Toward Better Language Use in Efficient Vision-Language-Action Models for Autonomous Driving

DGX agent

arXiv:2606.08684v1 Announce Type: new Abstract: We present BLUE, a minimal method for better language use in vision-language-action (VLA) models for autonomous driving (AD). Through extensive analysis

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models

DGX agent

arXiv:2503.08434v5 Announce Type: replace-cross Abstract: Recent advances in large-scale text-to-image models have revolutionized creative fields by generating visually captivating outputs from textua

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Breaking the Bubble: Asynchronous Pipeline Parallel Training with Bounded Weight Inconsistency

DGX agent

arXiv:2606.07881v1 Announce Type: new Abstract: Pipeline parallelism is essential for training large neural networks, but existing schedules trade off throughput, memory, and optimization consistency.

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Bridged SBI: Correcting Biased Low-Fidelity Posteriors for Cost-Efficient High-Fidelity Inference

DGX agent

arXiv:2606.09155v1 Announce Type: new Abstract: Accurate calibration of particle-based simulators is crucial for robotic earthwork simulation, but analytical calibration is challenging due to this tas

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

BSTabDiff: Block-Subunit Diffusion Priors for High-Dimensional Tabular Data Generation

DGX agent

arXiv:2606.09257v1 Announce Type: cross Abstract: High-Dimensional Low-Sample Size (HDLSS) tabular domains (e.g., omics) are characterized by n ll m, where n = number of samples, and m = number of fea

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

BUDDY: BUdget-Driven DYnamic Depth Routing for Adaptive Large Language Model Inference

DGX agent

arXiv:2606.09514v1 Announce Type: new Abstract: Large language models (LLMs) incur high inference cost due to their depth and parameter scale. Depth pruning can reduce latency by skipping redundant Tr

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

C3VD-DEFCOL: A Deformable Colonoscopy Dataset with Time-Resolved 3D Ground Truth and Realistic Appearance

DGX agent

arXiv:2606.07891v1 Announce Type: new Abstract: 3D reconstruction could improve colonoscopy by estimating mucosal coverage and alerting clinicians to missed regions during screening. However, algorith

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reasoning Models

DGX agent

arXiv:2606.08571v1 Announce Type: cross Abstract: Large language models frequently fail in a characteristic way: rather than acknowledging ignorance, they produce fluent but incorrect answers to quest

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

CamoSAM2: SAM2-oriented Prompt Auto-Refinement for Video Camouflaged Object Detection

DGX agent

arXiv:2504.00375v2 Announce Type: replace Abstract: The Segment Anything Model 2 (SAM2), a prompt-guided video foundation model, has remarkably performed in video object segmentation, drawing signific

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

DGX agent

arXiv:2505.11189v3 Announce Type: replace Abstract: Large language models (LLMs) can amplify misinformation, undermining societal goals such as the UN SDGs. We study three documented drivers of misinf

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Can we stabilize an inverted pendulum with feedback from a time-of-flight camera?

DGX agent

arXiv:2606.09237v1 Announce Type: new Abstract: Time-of-flight cameras are popular in robotics for providing direct depth information while being compact, inexpensive, and robust to lighting condition

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

Can You Trust What You See? Human and AI Detection of Synthetic Legal Evidence

DGX agent

arXiv:2606.07613v1 Announce Type: cross Abstract: Visual evidence has long been treated as a reliable form of legal proof, but advances in artificial intelligence (AI) are undermining that assumption.

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

CATPO: Critique-Augmented Tree Policy Optimization

DGX agent

arXiv:2606.08346v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving the reasoning capabilities of large language models

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Causal Agent Replay: Counterfactual Attribution for LLM-Agent Failures

DGX agent

arXiv:2606.08275v1 Announce Type: cross Abstract: When an LLM agent fails -- issues a refund it should not have, calls the wrong tool, leaks data -- existing tooling answers what happened (observabili

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Causal Longitudinal Prior-Fitted Networks for Counterfactual Outcome Prediction

DGX agent

arXiv:2606.05797v2 Announce Type: replace Abstract: Longitudinal treatment decisions from multivariate time-series data require predicting potential outcomes under future treatment sequences in the pr

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Chain of Flow: ECG-Conditioned 4D Cardiac Cine Generation from Patient-Specific Anatomical Anchor

DGX agent

arXiv:2602.22919v2 Announce Type: replace Abstract: Cardiac cine magnetic resonance imaging (MRI) is central to functional cardiac assessment, yet a full current cine sequence may not always be direct

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

CHIMERA-Bench: A Benchmark Dataset for Epitope-Specific Antibody Design

DGX agent

arXiv:2603.13431v3 Announce Type: replace-cross Abstract: Computational antibody design has seen rapid methodological progress, with dozens of deep generative methods proposed in the past three years,

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage Sites in China

DGX agent

arXiv:2606.08959v1 Announce Type: new Abstract: We introduce ChinaHeritaQA, a multimodal benchmark dataset for evaluating the cultural reasoning abilities of vision-language models (VLMs) on UNESCO Wo

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

CHROMA: Detecting AI-Generated Images through Inter-Channel Color-Space Correlations

DGX agent

arXiv:2606.08864v1 Announce Type: new Abstract: The rapid adoption of diffusion and large-scale generative models has made it increasingly challenging to distinguish synthetic imagery from real photog

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

ChronoPhyBench: Do MLLMs Truly Understand the World or Merely Exploit Language Priors?

DGX agent

arXiv:2606.07962v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated remarkable proficiency in open-world reasoning and understanding. Howe

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

CLASP: Language-Driven Robot Skill Selection and Composition using Task-Parameterized Learning

DGX agent

arXiv:2606.08169v1 Announce Type: cross Abstract: Enabling robots to understand and execute tasks from natural language commands while maintaining data efficiency remains challenging. Foundation model

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Claude Code-Driving Scenario Mining for the Argoverse 2 Challenge

DGX agent

arXiv:2606.09180v1 Announce Type: new Abstract: We present our submission to the CVPR 2026 Argoverse 2 Scenario Mining Challenge. Our system uses a four-stage pipeline: (1) autonomous code generation

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Closing the Sim-to-Real Gap: An Evaluation Framework for Autonomous Cyber Defense Configuration of Commercial EDR

DGX agent

arXiv:2606.08168v1 Announce Type: cross Abstract: Leading commercial endpoint detection and response (EDR) products have shifted from operator-configured rule sets to multi-component systems where aut

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Coarse-to-Fine Hierarchical Alignment for UAV-based Human Detection using Diffusion Models

DGX agent

arXiv:2512.13869v3 Announce Type: replace Abstract: Training object detectors demands extensive, task-specific annotations, yet this requirement becomes impractical in UAV-based human detection due to

model-releasesarxiv-cs-cv
9 Jun 2026
← Previous
1…140141142143144…361
Next →