AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning

DGX agent

arXiv:2607.26977v1 Announce Type: new Abstract: Travel planning is a demanding stress test for tool-using LLM agents: a usable itinerary is a single artifact that must be right along many axes at once

model-releasesarxiv-cs-cl
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Two Calls Beat Five Agents: Evaluating Multi-Agent Pipelines Against Self-Refinement for Local Language Models

DGX agent

arXiv:2607.26922v1 Announce Type: new Abstract: Multi-agent LLM pipeline systems break down the task among multiple roles for better reasoning, but are benchmarked mainly with large-scale commercial m

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Understanding Knowledge Transfer Mechanism in Heterogeneous MLLM Fusion: A Simple Linear Approach

DGX agent

arXiv:2607.26608v1 Announce Type: new Abstract: Training-free fusion of heterogeneous multimodal large language models (MLLMs) provides a direct route for cross-scale capability transfer, yet improvem

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

VEGA: Learning Navigation VLAs from In-the-Wild Egocentric Video with Geometric Trajectory Supervision

DGX agent

arXiv:2606.18426v2 Announce Type: replace Abstract: We introduce VEGA, an approach for training navigation VisionLanguage-Action (VLA) models from unlabeled egocentric navigation videos. Internet-scal

model-releasesarxiv-cs-ro
30 Jul 2026
Model Releases

Visual Credit Audit for Multimodal Spatial Reasoning

DGX agent

arXiv:2607.27069v1 Announce Type: new Abstract: Closed yes/no spatial benchmarks can reward a correct answer even when the image adds little support beyond no-image contexts. Under a fixed forced-choi

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

What Can Latent World Models Know? Physical Parameter Identifiability in Multimodal Predictive Representations

DGX agent

arXiv:2607.27017v1 Announce Type: new Abstract: A central premise of latent world models is that predicting the future forces a representation to internalize the physics of its environment. Which phys

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

When benchmark inferences do not compose: Projectibility in AI evaluation

DGX agent

arXiv:2607.26159v1 Announce Type: cross Abstract: An AI benchmark result rarely reaches a consequential claim in one step. Evaluators generalize it to further cases, interpret it as evidence of capabi

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

When Fish Look Alike: Tracking Identities with Dual-branch Elasticity

DGX agent

arXiv:2607.26412v1 Announce Type: new Abstract: Tracking dense, homogeneous targets like schooling fish remains a major challenge for multiple object tracking due to extreme inter-individual homogenei

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses

DGX agent

arXiv:2607.26348v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as synthetic users, stand-ins for human respondents whose simulated answers feed product, policy, and

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Where Detectors Fail: Closing the Tail-Domain Gap with Expert-Guided Mutual Distillation

DGX agent

arXiv:2607.26555v1 Announce Type: new Abstract: Multimodal fake news detectors often generalize poorly across domains because they learn to trust unreliable evidence: domain-specific shortcuts amplifi

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

WildShadowRemover: In-the-Wild Video Shadow Removal via Detail-Preserving Video Diffusion Models

DGX agent

arXiv:2607.26203v1 Announce Type: new Abstract: Video shadow removal in the wild remains challenging due to complex illumination, diverse shadow appearances, and limited training data. Despite its imp

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Zero-Fi: Zero-Shot Wi-Fi-Based Human Activity Recognition via Contrastive Signal-Language Alignment

DGX agent

arXiv:2607.26381v1 Announce Type: new Abstract: Wi-Fi-based human activity recognition has advanced substantially, but most existing methods assume a closed set of activities and require labeled Wi-Fi

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

A Cost-Effective Multimodal LLM Reasoning Framework for Question Answering over Irregular Clinical Time Series

DGX agent

arXiv:2607.25947v1 Announce Type: new Abstract: Question answering (QA) over irregular clinical time series (ICTS) plays a pivotal role in a wide range of healthcare applications. Although recent mult

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

A Physics-Informed Neural Operator for Thermal Ranking of Low-Cost Wall Materials in Hot-Dry Climates

DGX agent

arXiv:2607.25668v1 Announce Type: new Abstract: Identifying cost-effective indigenous building materials that minimise heat penetration through walls is critical for indoor thermal comfort in low-inco

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

A Unified Benchmark and Modality-Adaptive Network for Day-and-Night Drone-View Geo-Localization

DGX agent

arXiv:2607.25778v1 Announce Type: new Abstract: Most existing drone-view geo-localization (DVGL) benchmarks contain drone imagery captured under a single illumination condition and lack geographically

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Addressable Recall Compaction for Long Context-Window Control in AI Agents

DGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Agent Retrieval Bench: Evaluating Repository Context Retrieval for Coding Agents

DGX agent

arXiv:2607.24882v1 Announce Type: cross Abstract: Modern coding agents are usually evaluated by whether they eventually produce a correct patch, but patch generation depends on an earlier context-acqu

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

DGX agent

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in d

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

AIriskEval-edu Demo: Auditing of Pedagogical Risks in Educational Explanations

DGX agent

arXiv:2607.25634v1 Announce Type: new Abstract: We present AIriskEval-edu Demo, a platform that audits the pedagogical quality of instructional explanations and provides explainable audit results. The

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

AI's Capability in Assisting Scientific Research in Physics, Astrophysics, and Cosmology II: Project Planning and Proposal Evaluation

DGX agent

arXiv:2607.25881v1 Announce Type: new Abstract: We investigate how well large language models (LLMs) can assist scientific project planning and proposal evaluation. One-page project plans were indepen

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Aletheia: An Offline-First Clinical Decision Support System for Differential Diagnosis in Low-Resource Healthcare Settings

DGX agent

arXiv:2607.24814v1 Announce Type: new Abstract: Access to specialist clinical expertise remains severely limited across sub-Saharan Africa, where physician-to-patient ratios can fall below 1:25,000 in

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

AMPBench-MT: A Homology-Controlled Benchmark for Antimicrobial Peptide Potency, Spectrum, and Safety Prediction

DGX agent

arXiv:2607.25518v1 Announce Type: new Abstract: Computational AMP discovery is often evaluated through AMP/non-AMP recognition, yet follow-up decisions depend on assay-derived evidence such as target-

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

AnnoBench: A Benchmark for Visualization Annotation Generation

DGX agent

arXiv:2607.25911v1 Announce Type: cross Abstract: Annotation is among the most demanding visualization tasks to automate, as it simultaneously requires correctly navigating visual, semantic, and styli

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

At-the-Roofline Sparse Tensor Contractions on Vector Processors for Transformer Inference

DGX agent

arXiv:2607.25504v1 Announce Type: cross Abstract: Fine-grained weight pruning and activation sparsification have emerged as effective approaches for reducing the compute and memory cost of inference f

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Atmospheric Diffusion-Guided Spatio-Temporal Transformer for Nuclear Radiation Forecasting

DGX agent

arXiv:2607.24774v1 Announce Type: new Abstract: Nuclear radiation, the energy released during atomic decay, poses persistent risks to public health and the environment, and concerns have only grown si

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Authoring Agent Skills: A Software-Engineering Approach

DGX agent

arXiv:2607.25032v1 Announce Type: cross Abstract: Agent Skills are an emerging way to extend large language model agents with reusable procedural knowledge that the agent loads on demand. Anthropic in

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Automated Modernization of Machine Learning Engineering Notebooks for Reproducibility

DGX agent

arXiv:2602.07195v2 Announce Type: replace-cross Abstract: Interactive computational notebooks (e.g., Jupyter notebooks) are widely used in machine learning engineering (MLE) to program and share end-t

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

AVE-Compass: Towards Holistic Evaluation for Audio-Video Editing Abilities

DGX agent

arXiv:2607.24821v1 Announce Type: cross Abstract: While instruction-based video editing has advanced rapidly, real-world videos contain tightly coupled audio and visual signals, and editing one modali

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Balancing multiscale similarity and cartographic constraints: A similarity-driven optimization framework for line generalization

DGX agent

arXiv:2607.25474v1 Announce Type: new Abstract: Cartographic generalization is essential for generating multiscale map representations by balancing information preservation and cartographic readabilit

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Beyond Background Bias: Saliency-Driven Prototype Alignment for Dataset Distillation

DGX agent

arXiv:2607.25318v1 Announce Type: new Abstract: Dataset distillation aims to synthesize compact datasets that can approximate the performance of full-data training while significantly reducing computa

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Beyond Facial Consistency: Personalized Person Image Generation with Holistic Identity Preservation

DGX agent

arXiv:2607.25622v1 Announce Type: new Abstract: Personalized person image generation requires preserving subject identity across both local facial details and broader appearance cues. Existing methods

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Beyond Static Costs: Learning-Dynamics Aware Loss Functions for Long-Tailed Classification

DGX agent

arXiv:2607.25830v1 Announce Type: new Abstract: Deep learning models in computer vision face significant challenges when trained on long-tailed datasets, where a few majority classes dominate while ma

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Beyond 'What to Retrieve': Uncertainty in Retrieval-Augmented Code Generation

DGX agent

arXiv:2607.24884v1 Announce Type: cross Abstract: Repository-level code generation relies on heterogeneous evidence whose relevance, compatibility, and completeness are inherently uncertain. Similar-c

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Bits and Memories: Measuring Verbatim Extraction Across LLM Quantization

DGX agent

arXiv:2607.25451v1 Announce Type: new Abstract: Language models are almost always quantized before they are deployed, and a growing line of work asks whether quantization also lowers their privacy ris

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

Bridging Compute- and Data-Optimal Pretraining

DGX agent

arXiv:2607.25271v1 Announce Type: cross Abstract: Classical compute-optimal scaling laws assume an unbounded supply of fresh pretraining data, yet pretraining is increasingly entering a regime in whic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Building Large-Scale English-Romanian Literary Translation Resources with Open Models

DGX agent

arXiv:2509.07829v4 Announce Type: replace-cross Abstract: Literary translation has recently gained attention as a distinct and complex task in machine translation research, yet translation by small op

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CARE-MH: Towards Unified, Reproducible, and Comparable Evaluation of Mental Health LLMs

DGX agent

arXiv:2607.24754v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to provide mental health support, requiring reliable evaluation of safety, empathy, and therapeutic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CD-RMOT-Bench: Benchmarking the Cross-Domain Referring Multi-Object Tracking

DGX agent

arXiv:2607.25239v1 Announce Type: new Abstract: Referring multi-object tracking (RMOT) extends tracking from category-driven perception to language-guided understanding by grounding object trajectorie

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Certified-Gap Dual-Price Policies for Real-Time Truckload Bid Acceptance with Relocating, Clock-Constrained Resources

DGX agent

arXiv:2607.16891v2 Announce Type: replace-cross Abstract: A truckload carrier must accept or reject each load tender within seconds. The decision depends on fleet state, hours-of-service (HOS) clocks,

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Chart-Supported or Model-Supplied? Examining MLLM-Generated Claims for Accessible Visualization

DGX agent

arXiv:2607.25021v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can connect visualization patterns to external causes, consequences, and domain knowledge, but the evidential b

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition

DGX agent

arXiv:2607.25294v1 Announce Type: cross Abstract: Real-world tasks often require models to learn from task-specific context rather than relying only on pre-trained knowledge. While recent work has hig

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

COCO-OLAC: A Benchmark for Occluded Panoptic Segmentation and Image Understanding

DGX agent

arXiv:2409.12760v3 Announce Type: replace Abstract: To help address the occlusion problem in panoptic segmentation and image understanding, this paper proposes a new large-scale dataset named COCO-OLA

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models

DGX agent

arXiv:2607.24999v1 Announce Type: cross Abstract: LLM cognitive scores are increasingly summarized as per-ability profiles whose dimensions should converge across tasks, respond selectively to matched

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CogEEGAgent: Toward Autonomous Cognitive EEG Analysis with Grounded Execution and Selection-Aware Verification

DGX agent

arXiv:2607.25045v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis in cognitive studies requires specialized expertise and involves many defensible choices over contrasts, channels,

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference

DGX agent

arXiv:2607.25018v1 Announce Type: new Abstract: Large language model (LLM) cascades reduce inference cost by routing easy queries to a small model and deferring hard queries to a larger one. Productio

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

Cooperative Multi-UAV Navigation in Complex Environments via Systematic Multi-Agent Deep Reinforcement Learning

DGX agent

arXiv:2607.25754v1 Announce Type: new Abstract: Cooperative navigation of multi-agent UAVs in complex environments faces key challenges including local optima traps, sparse rewards, learning imbalance

model-releasesarxiv-cs-ro
29 Jul 2026
Model Releases

CoTinyVLA: Chain-of-Thought Distillation for a Sub-Billion-Parameter Vision-Language-Action Model

DGX agent

arXiv:2607.25487v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models translate natural-language commands into robot action sequences, but leading systems on the LIBERO-Plus robustness b

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

COVENANT: Natural-Language Workflow Compilation for Aligned Agent Execution

DGX agent

arXiv:2607.25400v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly entrusted with natural-language workflow instructions (e.g., retail-payment policies) that specify no

model-releasesarxiv-cs-ai
29 Jul 2026
← Previous
1…4546474849…357
Next →