AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

A time-series classification framework for individual-level absenteeism prediction under severe class imbalance

DGX agent

arXiv:2606.31532v1 Announce Type: new Abstract: Staff absenteeism imposes substantial operational costs in high-demand work environments such as healthcare, emergency services, meat processing, constr

model-releasesarxiv-cs-ai
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

A Transferable Learned Temporal Prior for Transmission Reconstruction and Decision-Relevant Uncertainty in Real Outbreak Labels

DGX agent

arXiv:2606.30842v1 Announce Type: new Abstract: Outbreak transmission reconstruction treats epidemiological timing and transmission labels as deterministic ground truth; neither has been systematicall

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Absorption-Feature-Guided Distance-Decoupled Estimation and Band Selection for LWIR Hyperspectral Passive Ranging

DGX agent

arXiv:2606.31824v1 Announce Type: new Abstract: Long-wave infrared (LWIR) hyperspectral observations contain distance-dependent atmospheric absorption signatures, providing a physical basis for long-r

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Accelerometry-Derived Digital Biomarkers for Cardiometabolic Risk: A Population-Representative Tabular Benchmark with Uncertainty Quantification

DGX agent

arXiv:2606.30702v1 Announce Type: cross Abstract: Structured tabular data dominates clinical medicine, yet existing benchmarks fail to reflect real-world properties like complex survey sampling, demog

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Adaptive Cluster-First Route-Second Decomposition for Industrial-Scale Vehicle Routing

DGX agent

arXiv:2606.31820v1 Announce Type: new Abstract: Large-scale capacitated vehicle routing problems (CVRPs) are commonly addressed using cluster-first route-second (CFRS) approaches that split a routing

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

AgentBound: Verifiable Behavioral Governance for Autonomous AI Agents

DGX agent

arXiv:2606.30970v1 Announce Type: new Abstract: Autonomous AI agents increasingly perform consequential actions on behalf of human principals, including financial transactions, external communications

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Agentic RAG-VLM: Affordance-Aware Retrieval-Augmented Generation with Self-Reflective Planning for Robotic Grasping

DGX agent

arXiv:2606.31200v1 Announce Type: new Abstract: Generalizable robotic grasping in cluttered environments is essential for deploying manipulators in unstructured human spaces, yet existing VLM-based me

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

AION: Aerial Indoor Object-Goal Navigation Using Dual-Policy Reinforcement Learning

DGX agent

arXiv:2601.15614v3 Announce Type: replace Abstract: Object-Goal Navigation (ObjectNav) requires an agent to autonomously explore an unknown environment and navigate toward target objects specified by

model-releasesarxiv-cs-ro
1 Jul 2026
Model Releases

An Empirical Study of Security Calibration in Large Language Models for Code

DGX agent

arXiv:2606.31159v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly transforming software development, yet their use in security-critical contexts raises a key question: do mode

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

An Executable Benchmarking Suite for Tool-Using Agents

DGX agent

arXiv:2605.11030v2 Announce Type: replace-cross Abstract: Closed-loop tool-using agents are increasingly evaluated in executable web, code, and micro-task environments, but benchmark reports often con

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Arena-T2I Hard: Benchmarking and Improving Faithfulness with Dependency-Aware Checklist

DGX agent

arXiv:2606.31711v1 Announce Type: new Abstract: Faithfulness -- how precisely a generated image aligns with its prompt -- is increasingly central to the real-world utility of text-to-image (T2I) model

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Artificial Intelligence in Sports: Insights from a Quantitative Survey among Sports Students in Germany about their Perceptions, Expectations, and Concerns regarding the Use of AI Tools

DGX agent

arXiv:2503.05785v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (AI) tools such as ChatGPT, Copilot, or Gemini have a crucial impact on academic research and teaching. Emp

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

AutoTrainess: Teaching Language Models to Improve Language Models Autonomously

DGX agent

arXiv:2606.31551v1 Announce Type: new Abstract: Training language models (LMs) remains a highly human-intensive process, even as frontier language model agents become increasingly capable at software

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

AxDafny: Agentic Verified Code Generation in Dafny

DGX agent

arXiv:2606.32007v1 Announce Type: new Abstract: We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

BayesBench: Evaluating LLM Belief Trajectories Under Multi-Turn Evidence Accumulation

DGX agent

arXiv:2606.30850v1 Announce Type: new Abstract: Large language models (LLMs) are typically deployed in multi-turn conversations, where each turn provides new evidence that should reduce epistemic unce

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Benchmarking Large Language Models on Floating-Point Error Classification

DGX agent

arXiv:2606.31308v1 Announce Type: new Abstract: This paper investigates the capability of Large Language Models (LLMs) to detect and classify floating-point errors statically in software code. We intr

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models

DGX agent

arXiv:2606.31338v1 Announce Type: cross Abstract: Recent music audio-language models achieve high accuracy on instrument question-answering benchmarks, but it remains unclear whether this reflects rob

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Beyond Clean Text: Evaluating Encoder and Decoder Robustness for Bangla Event Detection in Noisy Text

DGX agent

arXiv:2606.30914v1 Announce Type: new Abstract: Event detection (ED) systems are typically evaluated on clean, curated text, leaving their robustness to real-world noise largely unexplored, particular

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

DGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Beyond expert users: agents should help users construct preferences, not just elicit them

DGX agent

arXiv:2606.30863v1 Announce Type: new Abstract: Agents typically assume an expert user -- one with well-formed preferences about what they want -- and default to clarifying questions whenever the task

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Beyond Single Character: Evaluating MLLMs for Sentence-Level Oracle Bone Inscription Understanding

DGX agent

arXiv:2606.31169v1 Announce Type: new Abstract: Existing AI-assisted oracle bone inscription (OBI) visual recognition and understanding studies mainly focus on character-level, ignoring the long-form

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

BLUEX v2: Benchmarking LLMs on Open-Ended Questions from Brazilian University Entrance Exams

DGX agent

arXiv:2606.22723v2 Announce Type: replace Abstract: Although Large Language Models (LLMs) excel in many tasks, their assessment in Portuguese has received less attention, particularly for open-ended,

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Bridging Scientific Heritage: An Arabic--Russian Parallel Corpus and LLM Benchmark for Sustainable Knowledge Transfer

DGX agent

arXiv:2606.30943v1 Announce Type: new Abstract: Russian and Arabic are among the major languages of scientific communication. Language barriers impede the exchange of research results between these co

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Building an ASR Solution for Training and Assessing Children's Reading

DGX agent

arXiv:2606.31508v1 Announce Type: new Abstract: Automatic speech recognition for children's reading remains underdeveloped for most African languages, including Bambara, despite its potential value fo

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Calibrating the Evaluator: Does Probability Calibration Mitigate Preference Coupling in LLM Agent Feedback Loops?

DGX agent

arXiv:2606.31371v1 Announce Type: cross Abstract: When large language model (LLM) agents adapt their behavior through evaluator feedback, systematic evaluator biases propagate into the agent's learned

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Calibration, Not Compilation: Detecting and Repairing Misspecified Probabilistic Programs Written by Language Models

DGX agent

arXiv:2606.31630v1 Announce Type: new Abstract: Language models increasingly write probabilistic programs (in NumPyro, Stan, or Pyro), but a program that compiles, runs, and passes every unit test can

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Can Physician Expertise Improve Machine Learning Identification of Delirium?

DGX agent

arXiv:2606.30651v1 Announce Type: cross Abstract: Delirium is common in hospitalized patients and is often missed in routine care. We present a user-centered interactive machine learning (UC-iML) fram

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

CDR-Bench: Evaluating Faithful Execution of Compositional, Order-Sensitive Data Refinement Recipes

DGX agent

arXiv:2606.31435v1 Announce Type: new Abstract: Data refinement involves executing multi-step recipes over evolving text states, where both composition and execution order of processing operators dete

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

CHERRY: Compressed Hierarchical Experts with Recurrent Representational Yield

DGX agent

arXiv:2606.31796v1 Announce Type: cross Abstract: We study three complementary techniques for training compute-efficient language models. (1) Selective supervision and per-token efficiency. Selective

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Citation Discipline in Spec-Driven Development: A Cross-Model Empirical Study of Output Determinism and Automated Hallucination Detection in LLM-Generated Code

DGX agent

arXiv:2606.30689v1 Announce Type: cross Abstract: Spec-Driven Development (SDD) frameworks guide Large Language Model (LLM)-powered code generation through formal specifications, yet they differ funda

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

ClawArena-Team: Benchmarking Subagent Orchestration and Dynamic Workflows in Language-Model Agents

DGX agent

arXiv:2606.31174v1 Announce Type: new Abstract: Production large language-model (LLM) agents are increasingly deployed not as lone problem-solvers but as managers: a main model creates specialized sub

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Clinically Structured Rank-Gated LoRA for Cross-Benchmark Medical Question Answering

DGX agent

arXiv:2606.31432v1 Announce Type: new Abstract: Medical multiple-choice question answering requires parameter-efficient adaptation across heterogeneous knowledge domains and reasoning operations. A me

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Coarsening Bias from Variable Discretization in Causal Functionals

DGX agent

arXiv:2602.22083v2 Announce Type: replace-cross Abstract: Causal identification functionals often require integration over conditional densities of continuous variables, such as those arising in nonpa

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

CoLT: Teaching Multi-Modal Models to Think with Chain of Latent Thoughts

DGX agent

arXiv:2606.31986v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning has enabled multi-modal large language models (MLLMs) to tackle complex visual reasoning tasks by generating explicit i

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

ComAct: Reframing Professional Software Manipulation via COM-as-Action Paradigm

DGX agent

arXiv:2606.13239v2 Announce Type: replace-cross Abstract: Existing computer-use agents remain fundamentally limited in professional software manipulation: GUI-based agents suffer from fragile visual g

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Contextual Slate GLM Bandits with Limited Adaptivity

DGX agent

arXiv:2606.31449v1 Announce Type: new Abstract: We investigate the contextual slate bandit problem with generalized linear rewards under limited adaptivity. At each round, the learner is presented wit

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

CooperScene: Multi-Modal Cooperative Autonomy Benchmark with C-V2X Communication Characterization

DGX agent

arXiv:2606.31219v1 Announce Type: new Abstract: Cellular vehicle-to-everything (C-V2X) enables cooperative perception, prediction, and planning beyond the field of view of individual agents. However,

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Cross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on Romanian

DGX agent

arXiv:2606.31718v1 Announce Type: cross Abstract: Relation extraction (RE) for low-resource languages is typically constrained by the lack of annotated corpora. We investigate the feasibility of cross

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

CSTrader: A Testbed for Language-Grounded Trading in a Community-Driven Virtual Asset Market

DGX agent

arXiv:2606.31461v1 Announce Type: new Abstract: Niche asset markets, such as Counter-Strike 2 (CS2) weapon skins, are small, volatile, and heavily driven by community discussions and platform rules. T

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Curvature-Guided Module Localization for Low-Rank Detoxification of Backdoored Large Language Models

DGX agent

arXiv:2606.30899v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to large language models (LLMs) by causing otherwise benign systems to produce attacker-specified malicious beh

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

DANTE-W: Diffuse Albedo Neural Texturing in the Wild

DGX agent

arXiv:2606.30677v1 Announce Type: cross Abstract: Classical mesh texturing techniques blend captured multi-view images directly, which inevitably suffer from baked-in shading and casted shadows that c

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Dataset Construction for Training LLM to Learn Analog Circuit Knowledge

DGX agent

arXiv:2508.10409v3 Announce Type: replace-cross Abstract: This paper constructs a textual dataset for training large language models (LLMs) to learn analog circuit knowledge and customizes LLM trainin

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Diffusing Blame: Task-Dependent Credit Assignment in Biologically Plausible Dual-Stream Networks

DGX agent

arXiv:2606.31700v1 Announce Type: new Abstract: Biological neural circuits obey Dale's principle: each neuron's synapses are uniformly excitatory or inhibitory. Artificial networks that respect this c

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Disentangling Reasoning Logic to Resolve Explicit Knowledge Conflicts

DGX agent

arXiv:2508.01273v3 Announce Type: replace Abstract: Explicit knowledge conflicts, occurring when retrieved contexts contain contradictory information, pose a fundamental challenge for Large Language M

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Distill Once, Adapt Life-Long: Exploring Dataset Distillation for Continual Test-Time Adaptation

DGX agent

arXiv:2606.20196v2 Announce Type: replace Abstract: Continual Test-Time Adaptation (CTTA) aims to maintain model performance under evolving target domains by adapting online without labeled data. Howe

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Distilling Temporal Coherence into 2D Networks for Transrectal Ultrasound Prostate Video Segmentation

DGX agent

arXiv:2606.31198v1 Announce Type: cross Abstract: Real-time video segmentation of the prostate in Transrectal Ultrasound (TRUS) is essential for image-guided interventions. While conventional 2D metho

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Domain-Decomposed Randomized Neural Networks for Partial Differential Equations in Unbounded Domains

DGX agent

arXiv:2606.31342v1 Announce Type: cross Abstract: Partial differential equations on unbounded domains are challenging because the exterior region must be represented without excessive truncation error

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Dual Sparse Aggregation Transformer for Multispectral Object Detection

DGX agent

arXiv:2606.31015v1 Announce Type: new Abstract: Transformer-based approaches have obtained excellent performance in multispectral object detection tasks due to their ability to model long-range depend

model-releasesarxiv-cs-cv
1 Jul 2026
← Previous
1…99100101102103…361
Next →