AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
16,965 results
Model Releases

Transformers Struggle to Use Their Emergent World Models: Revisiting the Tower of Hanoi, and the Illusion of Thinking

DGX agent

arXiv:2608.07077v1 Announce Type: new Abstract: The Tower of Hanoi is a simple planning puzzle that in prior work has proven challenging for large reasoning models (LRMs). Current models solve the sta

model-releasesarxiv-cs-ai
10 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TransSLR: A Lightweight Transformer for Sign Language Recognition

DGX agent

arXiv:2608.06407v1 Announce Type: cross Abstract: Automated Sign Language Recognition for under-represented languages remains a largely unsolved problem. Central African Sign Language (CASL) exemplifi

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

TRIBE: Predicting Team Performance via Communication Behavior Ensembles

DGX agent

arXiv:2608.06926v1 Announce Type: new Abstract: Designing autonomous agents that effectively assist human teams hinges on understanding team dynamics, often without task specific knowledge. We present

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys

DGX agent

arXiv:2608.06404v1 Announce Type: new Abstract: Accurate 3D crop monitoring underpins data-driven precision agriculture by enabling field-scale analysis of plant structure, growth dynamics, and manage

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

UniREditBench: A Unified Reasoning-based Image Editing Benchmark

DGX agent

arXiv:2511.01295v3 Announce Type: replace Abstract: Recent advances in multi-modal generative models have driven substantial improvements in image editing. However, current generative models still str

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader

DGX agent

arXiv:2608.06474v1 Announce Type: new Abstract: Large language models increasingly generate complete websites from natural-language descriptions, and reinforcement learning has become a central approa

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

DGX agent

arXiv:2608.06704v1 Announce Type: new Abstract: Delegating a web task involves more than asking a question; it requires transferring a policy: what to verify, how to handle uncertainty, which preferen

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons

DGX agent

arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in tool READMEs and workshop papers, and the

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

WorldMark: A Plug-and-Play World Knowledge Interface for Cross-Host Language Model Watermarking

DGX agent

arXiv:2608.06416v1 Announce Type: cross Abstract: Watermarking traces the provenance of text produced by large language models by embedding statistically detectable signals during decoding. Existing s

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

YOLO-PEFT: Parameter-Efficient Fine-Tuning on YOLO Family

DGX agent

arXiv:2608.07051v1 Announce Type: new Abstract: Generic parameter-efficient fine-tuning (PEFT) methods transferred from language models can fail silently on real-time detectors, whose heterogeneous op

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination

DGX agent

arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores once memorized. extbf{Contamination mitigation

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

A Paragraph is Worth a Thousand Captions: Rethinking Text Supervision for Vision-Language Retrieval

DGX agent

arXiv:2608.05260v1 Announce Type: new Abstract: Contrastive vision-language models such as CLIP and BLIP are typically trained on short image captions, limiting their ability to retrieve images from d

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

A Six-Dimensional Taxonomy of Post-Training Adaptation Techniques with Applications in AI Governance

DGX agent

arXiv:2608.06246v1 Announce Type: new Abstract: Post-training adaptation has become central to modern machine learning practice and includes techniques such as retraining, fine-tuning, parameter-effic

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

A Two-Tier Perspective on Inference-Time Parallelism in Multi-Agent LLM Systems

DGX agent

arXiv:2608.05791v1 Announce Type: cross Abstract: Large language model (LLM)-driven multi-agent systems typically require multiple model invocations and complex coordination during inference, and thei

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

A Unified Risk View of Uncertainty: Posterior Risk for Disentanglement and Evaluation Beyond Proxies

DGX agent

arXiv:2608.05995v1 Announce Type: new Abstract: Reliable uncertainty estimates are critical in safety-sensitive applications, where understanding the sources of predictive uncertainty is essential. Th

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Abstract Event Causal Rules: Induction and Application

DGX agent

arXiv:2608.05205v1 Announce Type: new Abstract: Event-centric intelligent analytical systems heavily depend on explicit causal event knowledge for risk early warning, decision-making support and narra

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Adapting Vision Foundation Models with Cascaded Semantics

DGX agent

arXiv:2608.05393v1 Announce Type: new Abstract: Prompt tuning, a leading parameter-efficient adaptation paradigm in NLP, has recently been extended to computer vision. Visual prompt tuning (VPT) adapt

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Adaptive-WAM: Quality-Guided Early-Exit Planning from Intermediate Video-Diffusion Features

DGX agent

arXiv:2608.06008v1 Announce Type: new Abstract: Large video diffusion models provide rich spatiotemporal priors for autonomous driving, but existing world-action models often inherit the cost of itera

model-releasesarxiv-cs-ro
7 Aug 2026
Model Releases

Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation

DGX agent

arXiv:2503.03556v3 Announce Type: replace Abstract: Object affordance reasoning, the ability to infer object functionalities based on physical properties, is fundamental for task-oriented planning and

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

DGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

AI Playing Business Games: Benchmarking Large Language Models on Managerial Decision-Making in Dynamic Simulations

DGX agent

arXiv:2509.26331v2 Announce Type: replace Abstract: The rapid advancement of LLMs sparked significant interest in their potential to augment or automate managerial functions. One of the most recent tr

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Align-RAG: Alignment Is All You Need for TSFM In-Context Learning

DGX agent

arXiv:2608.05571v1 Announce Type: new Abstract: Retrieval-augmented forecasting promises to adapt frozen Time Series Foundation Models (TSFMs) to new domains without fine-tuning, but recent methods ty

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

An Axiomatic Benchmark for Evaluation of Scientific Novelty Metrics

DGX agent

arXiv:2604.15145v2 Announce Type: replace Abstract: The rigorous evaluation of the novelty of a scientific paper is, even for human scientists, a challenging task. With the increasing interest in AI s

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs

DGX agent

arXiv:2607.18056v2 Announce Type: replace Abstract: Frontier large language models (LLMs) are increasingly integrated into scientific workflows, yet their growing biological capabilities may outpace c

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New SheetSage-A2S Dataset

DGX agent

arXiv:2608.06165v1 Announce Type: cross Abstract: Existing audio-to-score (A2S) systems primarily focus on classical music, and the application to popular music remains underexplored. This paper first

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Autonomous Research Agents: A Survey of AI Scientists and the Verification Gap

DGX agent

arXiv:2608.05179v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used across the scientific research lifecycle: ideation, literature search, experiment design and e

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Benchmarking and Enhancing LLMs for Rule-Intensive Review of National Standard Documents

DGX agent

arXiv:2608.06312v1 Announce Type: new Abstract: Large language models (LLMs) increasingly support complex professional tasks, yet their capabilities in rule-intensive document review remain insufficie

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Benchmarking the Benchmarks: Evaluating Benchmarks for Conversational Agents

DGX agent

arXiv:2608.06329v1 Announce Type: cross Abstract: Task-oriented conversational agents are evaluated using curated or automatically generated benchmarks, yet benchmark quality is rarely assessed. Poor

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning

DGX agent

arXiv:2608.05250v1 Announce Type: new Abstract: Multi-task supervised fine-tuning (SFT) often casts a heterogeneous data mixture as a single optimization problem, even though different tasks may reach

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Beyond Rotations: AuroOFT for Expressive Quantized Orthogonal Fine-Tuning

DGX agent

arXiv:2608.05253v1 Announce Type: new Abstract: Quantized orthogonal fine-tuning (qoft) enables parameter-efficient adaptation of low-bit language models by learning structured activation rotations be

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Beyond Sequence Order: Syntax-Informed Positional Embeddings for Transformers

DGX agent

arXiv:2608.06111v1 Announce Type: cross Abstract: Positional embeddings (PE) in Transformers encode token distance and order but are largely agnostic to extit{syntactic structure}. We introduce extbf{

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Big, Bright, or Invisible: A Frozen-Feature Benchmark of 3D CT Foundation Models

DGX agent

arXiv:2608.05960v1 Announce Type: cross Abstract: Routine CT interpretation is inherently comprehensive, capturing incidental findings across the entire scan volume. 3D CT foundation models could assi

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

C^3PO: Evaluating Cross-Modal Composition and Counterfactual Performance in Omnimodal Models

DGX agent

arXiv:2608.05381v1 Announce Type: new Abstract: Current Multimodal Large Language Models (MLLMs) can process diverse sensory inputs, yet their reasoning remains heavily biased toward a dominant modali

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Can Open-Weight LLMs Produce Kernel-Verified Coq Proofs? A Pilot Study

DGX agent

arXiv:2608.05420v1 Announce Type: cross Abstract: Large language models (LLMs) can generate text that resembles a mathematical proof, but resemblance does not establish correctness. A formal proof che

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction

DGX agent

arXiv:2608.05359v1 Announce Type: new Abstract: CASCADE is an agentic framework that predicts downstream transcriptional effects of gene perturbation from precomputed ARACNe regulatory networks, expos

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Causal Episodic Memory for Feedback-Driven Agent Repair

DGX agent

arXiv:2608.05906v1 Announce Type: new Abstract: LLM agents that repair failures often discard successful corrections, forcing later episodes to rediscover similar solutions. We study whether finalized

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

DGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

ChronoVision: Temporal Reasoning via Latent State Reconstruction

DGX agent

arXiv:2608.05631v1 Announce Type: new Abstract: Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. T

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

CLARA: Clarification of Language Ambiguity through Result Analysis for Natural-Language Cancer Genomics Queries

DGX agent

arXiv:2608.05195v1 Announce Type: cross Abstract: A natural language interface can be used to make cancer genomics databases easier to use, but even if a question is perfectly fluent, its scientific m

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Clinician input steers AI toward accurate and harmful recommendations

DGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

CNM-BERT: A Drop-In Structural Embedding for Chinese Characters via Ideographic Description Sequences

DGX agent

arXiv:2608.05167v1 Announce Type: new Abstract: Token-based encoders like BERT treat Chinese characters as atomic identifiers, ignoring their recursive orthographic structure. Consequently, models rel

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

CodeGrep: An RL-Trained Retrieval Agent for LLM Coding Agents

DGX agent

arXiv:2608.05886v1 Announce Type: cross Abstract: Modern LLM coding agents such as Claude Code and OpenHands share a common inefficiency: they spend much of their token budget finding the file to patc

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning

DGX agent

arXiv:2608.05166v1 Announce Type: new Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our wo

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Continual Learning in Transition

DGX agent

arXiv:2608.06216v1 Announce Type: cross Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g.,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

ConWriter: Transition-Constrained Stateful Long-Form Story Generation with Lightweight Neuro-Symbolic Consistency Control

DGX agent

arXiv:2608.05169v1 Announce Type: new Abstract: Long-form story generation requires models to preserve narrative consistency across extended contexts, yet existing prompting-based methods often accumu

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

CREBench: Evaluating Large Language Models in Cryptographic Binary Reverse Engineering

DGX agent

arXiv:2604.03750v2 Announce Type: replace-cross Abstract: Reverse engineering (RE) is central to software security, particularly for cryptographic programs that handle sensitive data and are highly pr

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

CRINN: Contrastive Reinforcement Learning for Approximate Nearest Neighbor Search

DGX agent

arXiv:2508.02091v4 Announce Type: replace-cross Abstract: Approximate nearest-neighbor search (ANNS) algorithms have become increasingly critical for recent AI applications, particularly in retrieval-

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Cross-Architecture Steering Transfer in Language Models: A Systematic Empirical Study

DGX agent

arXiv:2608.05164v1 Announce Type: new Abstract: Independently trained large language models may develop shared internal representations of semantic concepts despite architectural differences -- but wh

model-releasesarxiv-cs-cl
7 Aug 2026
← Previous
1…1415161718…354
Next →