AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

AgenticECO: An Agentic Framework for ECO on 3D Integrated Circuits

DGX agent

arXiv:2608.03738v1 Announce Type: new Abstract: As Moore's law slows, the industry is turning to three-dimensional integration; yet in merged 3D-IC flows, routed designs expose bond-level defects with

model-releasesarxiv-cs-ai
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems

DGX agent

arXiv:2608.03744v1 Announce Type: new Abstract: Clinical decision support is moving toward committees of language-model agents deliberating on a shared workspace. We ask whether such committees can be

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

AI Agent Economics: Can Autonomous Economic Behavior Emerge among AI Agents under Minimal External Conditions?

DGX agent

arXiv:2608.03076v1 Announce Type: new Abstract: Multi-agent studies commonly place AI agents in predefined games, markets, or roles, making it difficult to distinguish endogenous economic organization

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

AI Security Leaderboard: Methodology, Results and Minimal Standard

DGX agent

arXiv:2608.03070v1 Announce Type: cross Abstract: Frontier AI model developers increasingly rely on layered safeguards to prevent catastrophic misuse, but little public evidence exists on how much pro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

AI World Cup 2026: Benchmarking Large Language Models for End-to-End Football Tournament Prediction

DGX agent

arXiv:2608.03416v1 Announce Type: new Abstract: Large language models (LLMs) are now regularly asked to forecast real-world events, but comparisons are often difficult because models receive different

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ANCHOR-RE: An Agentic Neuro-Symbolic Framework for Grounded Biomedical Relation Extraction

DGX agent

arXiv:2608.03154v1 Announce Type: new Abstract: Biomedical relation extraction (BioRE) extracts structured knowledge from biomedical literature for applications such as knowledge base construction and

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

ANNOTARES: A Dataset for Extracting Logical Structures from German Statutory Texts

DGX agent

arXiv:2608.03898v1 Announce Type: new Abstract: The automatic structural analysis of legal texts is a cornerstone of legal technology, yet the extraction of their logical components remains a signific

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Approximate Speculative Decoding

DGX agent

arXiv:2608.03447v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive generation by verifying a draft block with a target model in parallel. Under standard greedy verificat

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ArtECulture: Benchmarking Culture-Conditioned Visual Emotion Understanding in Multimodal Large Language Models

DGX agent

arXiv:2608.03358v1 Announce Type: new Abstract: Existing visual emotion understanding methods typically ignore cultural variations in emotional perception. We introduce culture-conditioned visual emot

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

AS-FedBridge: Pseudo-Spike Bridge Distillation for Heterogeneous ANN-SNN Federated Learning

DGX agent

arXiv:2608.03324v1 Announce Type: new Abstract: Federated learning enables collaborative model training across distributed edge devices while strictly preserving data privacy. To facilitate practical

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation

DGX agent

arXiv:2608.03990v1 Announce Type: new Abstract: Synthetic histopathology image generation has emerged as an approach that may address data scarcity in computational pathology, yet current evaluation m

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

ATFlash: Per-RoPE-Wavelength Attention Windows for Compute/Memory-Efficient LLM Inference

DGX agent

arXiv:2608.02947v1 Announce Type: cross Abstract: The attention score with rotary position embeddings (RoPE) decomposes exactly into a sum over its 2D-rotation frequency pairs, and each pair's wavelen

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization

DGX agent

arXiv:2502.11140v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become a cornerstone for automated visualization code generation, enabling users to create charts through na

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs

DGX agent

arXiv:2608.03450v1 Announce Type: cross Abstract: Reasoning in Multimodal Large Language Models (MLLMs) requires both fine-grained visual perception and rigorous logical deduction. Explicit text-based

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

BanglaWild: An In-the-Wild Bengali Scene Text Recognition Benchmark for OCR and Vision-Language Models

DGX agent

arXiv:2608.03884v1 Announce Type: cross Abstract: In-the-wild Bengali scene text recognition is largely unmeasured: existing resources target handwritten documents or constrained sign-board parsing, r

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

BBOWP-Bench: Evaluating LLMs on Black-Box Optimization Word Problems

DGX agent

arXiv:2608.02612v1 Announce Type: new Abstract: Formulating an optimization problem strongly affects the quality of the final solution, yet good formulations usually require substantial expertise. Rec

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms

DGX agent

arXiv:2505.14744v3 Announce Type: replace-cross Abstract: Traditionally, in Programming-by-example (PBE) the goal is to synthesize a program from a small set of input-output examples. Lately, PBE has

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Initialization Loss: A Systematic Study of Token Embedding Initialization Strategies for LLM Vocabulary Extension

DGX agent

arXiv:2608.03494v1 Announce Type: new Abstract: Vocabulary extension is an efficient way to adapt pretrained large language models (LLMs) to new languages, but the initialization of newly added token

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond Representational Similarity: Source-Conditioned Description-Length Gain for Generative Plagiarism Detection and Candidate Source Reranking

DGX agent

arXiv:2608.03859v1 Announce Type: cross Abstract: Large language models (LLMs) pose challenges to academic integrity and peer review. Yet generative plagiarism detection remains an underexplored and l

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety

DGX agent

arXiv:2601.17003v2 Announce Type: replace-cross Abstract: Mental-health AI safety is typically evaluated with small, simulation-based benchmarks that may not reflect the linguistic and contextual dive

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond the Gegenbauer Paradigm: q-Orthogonal Kernels for Machine Learning

DGX agent

arXiv:2608.03482v1 Announce Type: new Abstract: The performance of Support Vector Machines (SVMs) critically depends on the kernel function choice, which enables implicit mapping of data into high-dim

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Beyond the Single Camera: Agentic Multi-View Reasoning in Sports Video Understanding

DGX agent

arXiv:2607.11844v2 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) achieve strong performance on single-view video understanding benchmarks. However, sports videos inv

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Bridging Online and Offline Handwriting via Differentiable Physical Rendering

DGX agent

arXiv:2608.03198v1 Announce Type: new Abstract: Realistic handwritten text generation plays an important role in numerous applications, such as font design, biometric authentication, and robotic calli

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

BulkPR-Bench: Benchmarking Queue-Level Governance of Interacting Pull Requests

DGX agent

arXiv:2608.02685v1 Announce Type: cross Abstract: Coding-agent benchmarks increasingly cover long-horizon, end-to-end, and interactive development, but typically retain one requested outcome or a fixe

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

CADET: Physics-Grounded Causal Auditing and Training-Free Deconfounding of End-to-End Driving Planners

DGX agent

arXiv:2606.14438v3 Announce Type: replace-cross Abstract: End-to-end (E2E) autonomous-driving planners trained by imitation are prone to statistical shortcuts: they associate scene elements that merel

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss?

DGX agent

arXiv:2608.03983v1 Announce Type: cross Abstract: Optimizing compilers miss profitable transformations when their enabling semantics are absent from the analyzed program representation. We ask whether

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can LLM design high-quality experiments? A Comprehensive and Systematic Benchmark on Autonomous Experimental Design

DGX agent

arXiv:2608.03501v1 Announce Type: new Abstract: AI for Research (AI4Research) leverages AI to automate and improve scientific workflows. While experimental design is a critical stage of the research p

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can LLMs Test Terminal User Interfaces?

DGX agent

arXiv:2608.03743v1 Announce Type: cross Abstract: Terminal User Interfaces (TUIs) combine the stateful, screen-oriented behaviour of GUIs with terminal deployment and are now common in developer tools

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Can Text-to-Image Models Draw from the Right Frame of Reference?

DGX agent

arXiv:2608.03357v1 Announce Type: new Abstract: Spatial instruction following has become a crucial requirement for text-to-image (T2I) generation. A common challenge arises when directional expression

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

CARE-Bench: Benchmarking Patient-Facing LLM Triage

DGX agent

arXiv:2608.03731v1 Announce Type: new Abstract: Patient-facing medical LLMs and agents increasingly answer symptom questions before clinician contact, where the key safety question is what action the

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ChartAnno: Evaluating MLLMs for Chart Annotation Generation

DGX agent

arXiv:2608.03464v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have made significant progress in chart understanding, generation, and editing, but their ability to annotate e

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

ChiEngMixBench: Evaluating Large Language Models on Expert-Style Chinese-English Terminology Mixing

DGX agent

arXiv:2601.16217v2 Announce Type: replace-cross Abstract: Large language models increasingly mediate multilingual professional communication, where useful generation requires adapting to community con

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Confident but Unreliable: A Behavioral Safety Audit of Vision-Language Models on Brain MRI

DGX agent

arXiv:2608.02790v1 Announce Type: new Abstract: Vision-language models (VLMs), including medical specialists, are increasingly proposed for medical imaging, yet their stated confidence is rarely evalu

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

ConlangBench: Exploring Language Knowledge and Learning in LLMs through Diverse Constructed Languages

DGX agent

arXiv:2608.03505v1 Announce Type: new Abstract: Constructed languages (conlangs) are intentionally created human languages with a rich tradition of linguistic creativity. Despite their potential for s

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

CorePath: A Breast-Specialized Pathology Foundation Model for Core Needle Biopsy Diagnosis and Risk-Controlled Report Generation

DGX agent

arXiv:2608.03079v1 Announce Type: cross Abstract: Breast core needle biopsy (CNB) is central to breast cancer diagnosis yet remains challenging because limited tissue sampling, lesion heterogeneity, a

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Cross-Lingual Bias in Large Language Models: A Comparative Analysis of English and Swahili

DGX agent

arXiv:2608.03532v1 Announce Type: new Abstract: Large language models are increasingly deployed in multilingual contexts, yet safety alignment and bias evaluation remain overwhelmingly English-centric

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

CrossScope: A Role-Asymmetric World Model for Joint Dual-Scope Surgical Video Prediction

DGX agent

arXiv:2608.03211v1 Announce Type: new Abstract: Visual world models typically learn future dynamics from a single observation stream, limiting their ability to model cooperative systems with multiple

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

CUADebug: Diagnosing and Repairing Computer-Use Agent Failures

DGX agent

arXiv:2608.02643v1 Announce Type: cross Abstract: Computer-use agents (CUAs) operate real desktop and web interfaces through screenshots, mouse and keyboard actions, and stateful UI feedback, yet thei

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

CUDA MPC: A GPU-Native Solver for Model Predictive Control

DGX agent

arXiv:2608.03051v1 Announce Type: new Abstract: Model Predictive Control (MPC) delivers constraint-aware control, but its reliance on online optimization limits its use on systems with fast dynamics,

model-releasesarxiv-cs-ro
5 Aug 2026
Model Releases

Cura 1T: Specialized Model for Agentic Healthcare

DGX agent

arXiv:2607.15314v2 Announce Type: replace Abstract: Healthcare AI agents handle patient consultation, clinical reasoning over text and images, interactive diagnosis, and electronic health record (EHR)

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

DAIF: A Data-Driven Intermediate Fusion Framework for Multimodal Supervised Learning via Approximate Message Passing

DGX agent

arXiv:2608.02769v1 Announce Type: cross Abstract: Multimodal supervised learning seeks to leverage multiple heterogeneous data sources to improve predictive performance. A central challenge is determi

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces

DGX agent

arXiv:2608.03451v1 Announce Type: new Abstract: Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured file

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

DenialRAG: Single-Document RAG Poisoning via Embedded Parametric Denial

DGX agent

arXiv:2608.02678v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to corpus poisoning: an attacker who inserts a crafted document into the retrieval corpus

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Detecting Hallucinations and Recovering Verified Answers in Arabic Islamic Question Answering

DGX agent

arXiv:2608.03720v1 Announce Type: new Abstract: Large language models can generate fluent responses to Islamic questions while introducing factual errors that are difficult to identify. This paper pre

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

DiagChain: A Diagnostic Benchmark for Evaluating LLM Agents on Evidence-Grounded Attack Chain Reconstruction

DGX agent

arXiv:2608.03591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents offer a promising approach to attack chain reconstruction by retrieving and interpreting heterogeneous telemetry to

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Disentangling MLP Neuron Weights in Vocabulary Space

DGX agent

arXiv:2604.06005v2 Announce Type: replace Abstract: Interpreting the information encoded in language model weights remains a fundamental challenge in mechanistic interpretability. In this work, we int

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Distractor-Aware Truncation: Disentangling Context-Length Effects from Signal Loss in Long-Context LLM Benchmarks

DGX agent

arXiv:2608.03297v1 Announce Type: new Abstract: A standard claim in the literature on retrieval-augmented and memory-augmented language models is that shorter context is better when the relevant infor

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Diversity is Not Ambiguity: Toward Accurate and Efficient Ambiguity Detection for Open-Domain QA

DGX agent

arXiv:2608.03177v1 Announce Type: new Abstract: How can question answering (QA) systems determine whether a query is ambiguous? Ambiguity detection is essential in open-domain QA, as misclassification

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…2526272829…357
Next →