AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Linguistic Relative Policy Optimization for Video Anomaly Reasoning

DGX agent

arXiv:2607.00654v1 Announce Type: new Abstract: Video anomaly detection (VAD) with multimodal large language models has shown strong potential, yet most existing methods still depend on large-scale an

model-releasesarxiv-cs-cv
2 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Linkify: Learning from Interface-Augmented Assembly Graphs

DGX agent

arXiv:2607.01205v1 Announce Type: new Abstract: We present Linkify, a framework for learning from interface-augmented assembly graphs to enable context-aware part retrieval in mechanical assemblies. W

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

LLM-Guided ODE Discovery and Parameter Inference from Small-Cohort Aggregate Data

DGX agent

arXiv:2607.00733v1 Announce Type: cross Abstract: Mechanistic modeling via ordinary differential equations (ODEs) provides interpretable descriptions of complex dynamics and enables inference of under

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

LLVM-Bench: Benchmarking and Advancing Large Language Models for LLVM Compiler Issue Resolution

DGX agent

arXiv:2607.00700v1 Announce Type: cross Abstract: LLVM is a widely used compiler infrastructure whose scale and complexity make issue resolution labor-intensive and challenging. Although large languag

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Logit-Contribution Scoring Identifies Non-Literal Retrieval Heads

DGX agent

arXiv:2607.01002v1 Announce Type: cross Abstract: In long-context use, large language models frequently synthesize answers from the meaning of a relevant context span rather than literally copy-pastin

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

LongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language Models

DGX agent

arXiv:2607.01086v1 Announce Type: cross Abstract: The evaluation of long-term video quality understanding remains an open challenge for large vision-language models (LVLMs). Existing video quality ben

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Lost in the Tail: Addressing Geographic Imbalance in Urban Visual Place Recognition

DGX agent

arXiv:2607.00090v1 Announce Type: cross Abstract: Urban-scale Visual Place Recognition (VPR) aims to identify the geographic location of a query image by matching it against a geo-tagged database. Whi

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

LUMA: Benchmarking Segmentation via a Lightweight Universal Mask Adapter

DGX agent

arXiv:2607.00687v1 Announce Type: cross Abstract: Comparing transformer backbones for image segmentation is confounded: each is paired with a different decoder, recipe, and pretraining, so reported di

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

LuxIT: A Luxembourgish Instruction Tuning Dataset from Monolingual Seed Data

DGX agent

arXiv:2510.24434v3 Announce Type: replace Abstract: The effectiveness of instruction-tuned Large Language Models (LLMs) is often limited in low-resource linguistic settings due to a lack of high-quali

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

LV-ROVER: Multi-Stream Tesseract Voting for Maltese Paragraph OCR

DGX agent

arXiv:2607.00250v1 Announce Type: new Abstract: Maltese has decent text corpora and pretrained language models, but, like many languages outside the handful with large OCR benchmarks, only a single kn

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Mapping the Evaluation Frontier: An Empirical Survey of the Bias-Reliability Tradeoff Across Eleven Evaluator-Agent Conditions

DGX agent

arXiv:2607.00304v1 Announce Type: cross Abstract: The bias-reliability tradeoff conjectures that LLM evaluation systems are constrained in (gamma, H, CV) space, where evaluator coupling (gamma), strat

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

MemSyco-Bench: Benchmarking Sycophancy in Agent Memory

DGX agent

arXiv:2607.01071v1 Announce Type: cross Abstract: Memory has emerged as a cornerstone of modern LLM-based agents, supporting their evolution from single-turn assistants to long-term collaborators. How

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

MEPA: Multi-Scale Representation Alignment for Visual Autoregressive Modeling with Mixture of Experts

DGX agent

arXiv:2607.00371v1 Announce Type: cross Abstract: Visual AutoRegressive modeling (VAR) has pioneered a coarse-to-fine multi-scale autoregressive generative paradigm, demonstrating strong capabilities

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

MindAU: EEG-Conditioned Facial Action Unit Editing via Dual-Stream Manifold Alignment

DGX agent

arXiv:2607.00410v1 Announce Type: new Abstract: Recent brain decoding studies have made substantial progress in reconstructing externally perceived visual content from neural signals. However, using e

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

MindEdit-Bench: Benchmarking Object-Level Counterfactual Spatial Reasoning in VLMs from In-the-Wild Photos

DGX agent

arXiv:2607.00491v1 Announce Type: cross Abstract: Benchmarks for vision-language models (VLMs) mostly test observational spatial reasoning: models describe relations already visible in the input. Exis

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation

DGX agent

arXiv:2602.21397v2 Announce Type: replace Abstract: Prompt learning has become a dominant paradigm for adapting vision-language models (VLMs) such as CLIP to downstream tasks without modifying pretrai

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models

DGX agent

arXiv:2607.01117v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) have shown strong progress in video understanding, yet they still suffer from hallucinations that are inconsiste

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

MolSafeEval: A Benchmark for Uncovering Safety Risks in AI-Generated Molecules

DGX agent

arXiv:2607.00464v1 Announce Type: cross Abstract: Current molecular generation benchmarks emphasize task complexity, molecule novelty, and property alignment; they largely overlook a critical concern:

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

MSQA: A Natively Sourced Multilingual and Multicultural SimpleQA Benchmark

DGX agent

arXiv:2607.00724v1 Announce Type: new Abstract: Multilingual fluency often invites a stronger assumption: a model that can speak a user's language must also understand the culture encoded by that lang

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Multi-Hypothesis Test-Time Adaptation to Mitigate Underspecification

DGX agent

arXiv:2607.00259v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) seeks to improve model robustness under distribution shifts by adapting parameters using unlabeled target data. However, in

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Multi-Label Node Classification with Label Influence Propagation

DGX agent

arXiv:2607.00671v1 Announce Type: cross Abstract: Graphs are a complex and versatile data structure used across various domains, with possibly multi-label nodes playing a particularly crucial role. Ex

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning

DGX agent

arXiv:2607.00461v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are often constrained by a language-space bottleneck, forcing complex visual reasoning into discrete tokens whi

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

MultiSynt/MT: Trillion-Token Multi-Parallel Pre-Training Data Translated Across 36 Languages

DGX agent

arXiv:2607.00890v1 Announce Type: new Abstract: Open web-scale pre-training corpora remain concentrated in English, limiting multilingual LLM development. We introduce MultiSynt/MT, an open synthetic

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Neural Network-Based Estimation of Time-Dependent Parameters in AR(p) Processes

DGX agent

arXiv:2607.00470v1 Announce Type: cross Abstract: We investigate a forecasting framework based on a simple discrete-time dynamic model with coefficients varying in time. The parameters of the model ar

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Next-Frame Decoding for Ultra-Low-Bitrate Image Compression with Video Diffusion Priors

DGX agent

arXiv:2603.15129v3 Announce Type: replace Abstract: We present a novel paradigm for ultra-low-bitrate image compression (ULB-IC) that exploits the ``temporal'' evolution in generative image compressio

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Not All Prediction Targets Keep Training-Free Diffusion Guidance on the Manifold

DGX agent

arXiv:2607.00647v1 Announce Type: new Abstract: Training-free guidance (TFG) steers a pretrained diffusion model toward a desired attribute at inference. To be effective, this guidance must be applied

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

OmniFall: From Staged Through Synthetic to Wild, A Unified Multi-Domain Dataset for Robust Fall Detection

DGX agent

arXiv:2505.19889v3 Announce Type: replace Abstract: Visual fall detection models are usually trained on small, staged datasets. Their real-world utility remains unclear; such data lacks diversity and

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale

DGX agent

arXiv:2602.05711v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) architectures are evolving towards finer granularity to improve parameter efficiency. However, existing MoE designs f

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

On the Reliability of Cue Conflict and Beyond

DGX agent

arXiv:2603.10834v4 Announce Type: replace-cross Abstract: Understanding how neural networks rely on visual cues offers a human-interpretable view of their internal decision processes. The cue-conflict

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

OSCAR: Occupancy-based Shape Completion via Acoustic Neural Implicit Representations

DGX agent

arXiv:2603.08279v2 Announce Type: replace Abstract: Accurate 3D reconstruction of vertebral anatomy from ultrasound is important for guiding minimally invasive spine interventions, but it remains chal

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Pano2World: End-to-End 3D Generation via Unified Multi-View Sequences

DGX agent

arXiv:2607.00832v1 Announce Type: cross Abstract: A single panorama captures the full visual sphere from one camera center, yet confines users to looking around in place without enabling true scene ex

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Persona Without Substrate: Regime-Dependence and the LLM Individuation Problem

DGX agent

arXiv:2607.00006v1 Announce Type: cross Abstract: Beckmann & Butlin's (2026) ontological framework for the LLM individuation problem inherits an unargued cross-regime co-reference assumption from the

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

PHREEQC-MCQ-200: A Diagnostic Benchmark for Tool-Augmented Scientific Simulator Agents

DGX agent

arXiv:2607.00436v1 Announce Type: new Abstract: Large language model agents are increasingly connected to scientific software, yet it remains unclear when tool access makes scientific computation more

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking

DGX agent

arXiv:2607.00115v1 Announce Type: new Abstract: This paper explores multi-turn visual reasoning and observes that MLLMs repeatedly fail to localize the target, leading to long, redundant trajectories.

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

PorTEXTO: A European Portuguese Benchmark for Visual Text Extraction

DGX agent

arXiv:2606.19096v2 Announce Type: replace Abstract: European Portuguese (pt-PT) is largely absent from OCR benchmarks, which skew toward high-resource languages. The few benchmarks that cover pt-PT fo

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Post-Training Pruning for Diffusion Transformers

DGX agent

arXiv:2607.00927v1 Announce Type: cross Abstract: Diffusion Transformers (DiTs) have demonstrated impressive performance in image generation but suffer from substantial computational overhead and reso

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study

DGX agent

arXiv:2607.00049v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in Agile Software Development for documentation, coaching, and training. As practitioners adopt the

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

QuaMoE-DRF: Proactive Beam and Rate Adaptation via Multimodal Dynamic Radio Map Forecasting in ISAC Networks

DGX agent

arXiv:2607.00974v1 Announce Type: cross Abstract: Static radio maps provide location-dependent propagation priors, but they cannot capture short-term blockage caused by moving objects. Direct sensing-

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Quantifying the Affective Gap: A Zero-Shot Evaluation of LLMs on Fine-Grained Emotion Taxonomies

DGX agent

arXiv:2607.00968v1 Announce Type: new Abstract: Emotion recognition in natural language is a foundational challenge in affective computing, with critical implications for human-computer interaction, m

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Quantum vs. Classical Machine Learning: A Unified Empirical Comparison

DGX agent

arXiv:2607.01197v1 Announce Type: new Abstract: Quantum computing has emerged as a promising computational paradigm for machine learning (ML), with the potential to offer computational advantages over

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Radial Interaction Tomography: Recognizing Non-Transitive Evolutionary Games from One Range-Expansion Image

DGX agent

arXiv:2607.00378v1 Announce Type: new Abstract: Colored sectors in a microbial range expansion encode more than lineage survival counts. We formulate a computer-vision inverse problem: from one endpoi

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

RC-GeoCP: Geometric Consensus for Radar-Camera Collaborative Perception

DGX agent

arXiv:2603.00654v3 Announce Type: replace Abstract: Collaborative perception (CP) improves scene understanding through multi-agent information sharing, yet LiDAR-centric systems remain costly and vuln

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models

DGX agent

arXiv:2607.00852v1 Announce Type: cross Abstract: This work studies the hidden-state inversion problem: recovering the original input token sequence of a decoder-only language model from its last-laye

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail

DGX agent

arXiv:2607.00310v1 Announce Type: cross Abstract: Foundation video diffusion models are increasingly viewed as world simulators for embodied agents, yet their pretraining on internet-scale generic vid

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

DGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations

DGX agent

arXiv:2607.01181v1 Announce Type: cross Abstract: RL with verifiable rewards (RLVR) has emerged as a powerful paradigm for training LMs on tasks with well-defined success metrics, such as code generat

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios

DGX agent

arXiv:2511.18011v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated powerful capabilities in general spatial understanding and reasoning. However, their fine

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Robust 3D Alignment of Generative Reconstructions via Partial Monocular Observations

DGX agent

arXiv:2607.00498v1 Announce Type: new Abstract: Aligning generative 3D reconstructions with partial monocular observations is a critical but under-explored challenge in computer vision. This task is i

model-releasesarxiv-cs-cv
2 Jul 2026
← Previous
1…979899100101…361
Next →