AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

ARDY: Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation

DGX agent

arXiv:2607.08741v1 Announce Type: cross Abstract: Generating realistic 3D human motions in real-time within interactive applications is key for animation, simulation, and humanoid robotics. While rece

model-releasesarxiv-cs-cv
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Attribute Retrieving for Open-Vocabulary Endoscopic Compositional Referring Segmentation

DGX agent

arXiv:2607.08397v1 Announce Type: new Abstract: Referring Image Segmentation (RIS) aims to segment image regions specified by natural language, enabling fine-grained and controllable visual understand

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

DGX agent

arXiv:2607.08745v1 Announce Type: new Abstract: Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as sc

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Bayesian Deep Learning for Discrete Choice

DGX agent

arXiv:2505.18077v3 Announce Type: replace-cross Abstract: Discrete choice models (DCMs) are used to analyze individual decision-making in contexts such as transportation choices, political elections,

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Benchmark Evaluation of Feredated Learning on Multi-organ Images

DGX agent

arXiv:2607.08219v1 Announce Type: new Abstract: The privacy requirements of medical data and its substantial variations across organs and modalities hinder the clinical implementation of medical AI. F

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Beyond Backpropagation: Monte Carlo Method Can Train Deep Neural Networks

DGX agent

arXiv:2607.08406v1 Announce Type: new Abstract: Backpropagation (BP) dominates deep learning training, but its reliance on gradients brings inherent troubles -- vanishing and exploding gradients. The

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Beyond Thermal Imaging: Inferring Thermophysical Properties from Time-Resolved Thermal Observations

DGX agent

arXiv:2607.07962v1 Announce Type: cross Abstract: Inferring latent physical properties from sensory observations is a fundamental challenge in machine perception. Among available sensing modalities, t

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Beyond wheelchairs and blindfolds: Investigating disability stereotypes in T2I models with INCLUDE-BENCH

DGX agent

arXiv:2607.08515v1 Announce Type: new Abstract: Text-to-image (T2I) models have been shown to exhibit social biases. Prior work has mainly focused on gender, skin tone, and cultural representation wit

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

BiasBench: A reproducible benchmark for tuning the biases of event cameras

DGX agent

arXiv:2504.18235v2 Announce Type: replace Abstract: Event-based cameras are bio-inspired sensors that detect light changes asynchronously for each pixel. They are increasingly used in fields like comp

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Blind-Spots-Bench: Evaluating Blind Spots in Multimodal Models

DGX agent

arXiv:2607.08317v1 Announce Type: new Abstract: Modern AI models achieve strong performance on many established benchmarks, yet they still fail on tasks that humans find almost trivial, such as manipu

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

CausalDS: Benchmarking Causal Reasoning in Data-Science Agents

DGX agent

arXiv:2607.08093v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with advanced tool use. Yet the relevant b

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Classical versus Deep Mirror-Symmetry Scoring: A Benchmark of Thirteen Methods

DGX agent

arXiv:2607.08379v1 Announce Type: new Abstract: Quantifying how mirror-symmetric an image is about a given axis (symmetry scoring) underpins applications from visual aesthetics to medical imaging, yet

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Closing the Null Space: Guidance-Aware Quantization for Classifier-Free Diffusion

DGX agent

arXiv:2607.08241v1 Announce Type: new Abstract: Deploying classifier-free guidance (CFG) diffusion models under real-world compute budgets requires quantization, yet existing post-training quantizatio

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

COALA: Robust Contextualized Speech-augmented Language Modeling for ASR via Contrastive Regularizer and Biasing Score Estimation

DGX agent

arXiv:2607.08117v1 Announce Type: new Abstract: Contextual biasing seeks to integrate external knowledge into automatic speech recognition (ASR) systems to accurately recognize domain-specific entitie

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editing

DGX agent

arXiv:2607.08497v1 Announce Type: cross Abstract: Recent unified multimodal models show a single architecture can jointly perform vision/language understanding and image generation/editing. However, t

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Collective Intelligence with Foundation Models

DGX agent

arXiv:2607.07729v1 Announce Type: cross Abstract: As foundation models grow in scale and diversity, coordinating multiple models into cooperative reasoning systems offers a path toward safer, more rel

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Compete Then Collaborate: Frontier AI Teachers Build a Verifiable Curriculum to Improve a Coding Student Beyond Imitation

DGX agent

arXiv:2607.08255v1 Announce Type: new Abstract: Large language models increasingly serve as teachers generating training data for smaller students. Prior multi-teacher knowledge distillation methods m

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Concretized Proposition Prompting Resolves Composition-Knowledge Dichotomy in Large Language Models

DGX agent

arXiv:2607.08018v1 Announce Type: new Abstract: LLMs often struggle to balance compositionality with knowledgeability, a challenge we define as Composition-Knowledge Dichotomy. To address this, we pro

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Context Graphs for Proactive Enterprise Agents

DGX agent

arXiv:2607.07721v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) and agentic frameworks have advanced enterprise AI considerably, yet agents remain fundamentally reactive: they wai

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Continual Test-Time Adaptation in Computer Vision: Methods, Benchmarks, and Future Directions

DGX agent

arXiv:2607.08164v1 Announce Type: new Abstract: Deep neural nets achieve remarkable performance when training and test data share the same distribution, but this assumption frequently breaks in real-w

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Cross-seed explainability using Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoders

DGX agent

arXiv:2607.08499v1 Announce Type: new Abstract: We present a Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoder (SAE) for extracting cross-seed universal features from independently trai

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization

DGX agent

arXiv:2603.00910v2 Announce Type: replace-cross Abstract: Layer-wise capacity in large language models is highly non-uniform: some layers contribute disproportionately to loss reduction, whereas other

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

DeepSWE: Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasks

DGX agent

arXiv:2607.07946v1 Announce Type: cross Abstract: DeepSWE is a benchmark of 113 original, long-horizon software engineering tasks for evaluating coding agents. Most public agentic coding benchmarks fo

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

DeltaV: Thinking with Visual State Updates in Unified Large Multimodal Models

DGX agent

arXiv:2607.08434v1 Announce Type: new Abstract: Current Unified Large Multimodal Models (ULMMs) support interleaved multimodal reasoning through textual reasoning and intermediate visual states, but t

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

DexVerse: A Modular Benchmark for Multi-Task, Multi-Embodiment Dexterous Manipulation

DGX agent

arXiv:2607.08751v1 Announce Type: new Abstract: Building general-purpose dexterous manipulation policies requires benchmarks that go beyond isolated tasks to systematically evaluate policies across di

model-releasesarxiv-cs-ro
10 Jul 2026
Model Releases

Diagnosing Shape-Prior Shortcuts in Long-Range Single-Shot Fringe Projection Profilometry

DGX agent

arXiv:2606.17093v2 Announce Type: replace Abstract: Learning-based single-shot fringe projection profilometry (FPP) has been studied almost entirely at close range, and the networks used are evaluated

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Diarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech

DGX agent

arXiv:2607.08208v1 Announce Type: new Abstract: This paper describes our self-designed system for Task 1 of the MLC-SLM 2026 Challenge for multilingual two-speaker conversational speech. The system co

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Different Teachers, Different Capabilities: Sub-1B On-Device Distillation for Structured Text Enrichment

DGX agent

arXiv:2607.08268v1 Announce Type: new Abstract: High-volume structured extraction pays a large model's latency on every item, so distilling the task into a small on-device model is attractive: compara

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Dive Into the Implicit Biases of Low-rank Vision-language Alignment

DGX agent

arXiv:2607.08194v1 Announce Type: new Abstract: Vision-language alignment, the stage that bridges pretrained vision encoders and large language models, is widely treated as a form of pretraining requi

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Do Transformations Reveal the Truth? Generative Residual Learning for Generalized AI-Generated Image Detection

DGX agent

arXiv:2607.08674v1 Announce Type: new Abstract: The rapid advancement of generative AI has enabled the creation of highly realistic deepfake media, posing significant threats, including misinformation

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Do You Need a Frontier Model as a Citation Verifier? Benchmarking Rubric LLMs for Deep-Research Source Attribution

DGX agent

arXiv:2607.08700v1 Announce Type: new Abstract: Reinforcement learning increasingly relies on an LLM judge to score each rubric criterion, and that judge acts as the reward model during training. Befo

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

DominoTree: Conditional Tree-Structured Drafting with Domino for Speculative Decoding

DGX agent

arXiv:2607.08642v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting several tokens and verifying them in parallel. Block-diffusion drafters such as DFlash produc

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Dual-Correlation Hypergraph Network for Unaligned RGBT Video Object Detection and A Large-scale Benchmark

DGX agent

arXiv:2607.08191v1 Announce Type: new Abstract: RGB-Thermal (RGBT) Video Object Detection (VOD) has gained significant traction due to its ability to overcome the limitations of conventional RGB-based

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Echoes Across Vietnam's Highlands, Delta, and Coast: A Multilingual Corpus for Cham, Khmer, and Tay-Nung

DGX agent

arXiv:2607.08362v1 Announce Type: new Abstract: Vietnam's ethnic minority languages are almost absent from the field of Natural Language Processing (NLP), and the challenge goes beyond data scarcity:

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Equivariant Quantum Clustering with Differential Privacy: Parameter-Efficient Privacy-Preserving Analysis Across Heterogeneous Sensitive Datasets

DGX agent

arXiv:2607.08092v1 Announce Type: cross Abstract: Privacy-preserving clustering is critical for analyzing sensitive data in healthcare, cybersecurity, and enterprise applications, where maintaining da

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Evaluating the Generalizability of Foundation Models for Extreme Environmental Events: Case Study of California Wildfire PM2.5

DGX agent

arXiv:2607.07951v1 Announce Type: new Abstract: Wildfire smoke events produce extreme PM_{2.5} concentrations that pose severe public health risks, yet forecasting rare, hazardous-level spikes remains

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

FabriVLA: A Lightweight Vision-Language-Action Model for Precise Multi-Task Manipulation

DGX agent

arXiv:2607.08575v1 Announce Type: new Abstract: We present FabriVLA, a lightweight Vision-Language-Action model for Precise Multi-Task Manipulation. FabriVLA combines an InternVL3.5 vision-language ba

model-releasesarxiv-cs-ro
10 Jul 2026
Model Releases

False Confidence: Automated Labels Confound Fairness Audits in Cervical Spine Segmentation

DGX agent

arXiv:2607.07852v1 Announce Type: cross Abstract: Automated segmentation of cervical-spine MRI is increasingly used in clinical workflows, yet no fairness audit exists for this anatomy. We show that a

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Formal Mechanisms for Market Stability in Self-Interested Agent Societies: A Marketplace Simulation Study

DGX agent

arXiv:2607.08652v1 Announce Type: new Abstract: Self-interested agents, left unconstrained, tend toward defection in repeated social dilemmas, causing cooperative gains from trade to collapse. This pa

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Functional and Secure Code Generation with Task Vectors

DGX agent

arXiv:2607.07881v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for code generation, but they struggle to generate functional code free of security vulnerabilities

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Generalization Theory for Through-the-Wall Radar Human Activity Recognition

DGX agent

arXiv:2607.08144v1 Announce Type: cross Abstract: Through-the-wall radar (TWR) human activity recognition (HAR) is important for non-line-of-sight indoor sensing, security monitoring, and emergency re

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos

DGX agent

arXiv:2512.01803v3 Announce Type: replace Abstract: Despite rapid advances in video generative models, robust metrics for evaluating visual and temporal correctness of complex human actions remain elu

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Geometry and Gradient-based Partitioning for Panoramic Outdoor Reconstruction

DGX agent

arXiv:2607.08769v1 Announce Type: new Abstract: Scaling 3D Gaussian Splatting (3DGS) to large outdoor scenes is costly in both data acquisition and computation. Adopting panoramic images with equirect

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Graph-Regularized Deep Learning for EEG-Based Emotion Recognition with Psychologically-Grounded Label Structure

DGX agent

arXiv:2607.07773v1 Announce Type: cross Abstract: EEG-based emotion recognition is critical for mental health monitoring and affective brain-computer interfaces, yet existing deep learning approaches

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Hallucination Self-Play: Bootstrapping Reinforced Detector via Evolved Generator

DGX agent

arXiv:2607.07993v1 Announce Type: new Abstract: Identifying faithfulness hallucinations in LLM-generated outputs remains challenging due to the scarcity of high-quality annotated data. Recent work rel

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales

DGX agent

arXiv:2607.08705v1 Announce Type: new Abstract: Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, posing unp

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

ICDAR 2026 HIPE-OCRepair Competition on LLM-Assisted OCR Post-Correction for Historical Documents

DGX agent

arXiv:2607.08143v1 Announce Type: cross Abstract: We present the results of HIPE-OCRepair-2026, an ICDAR competition on LLM-assisted OCR post-correction of historical documents. OCR post-correction re

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation

DGX agent

arXiv:2607.08758v1 Announce Type: new Abstract: Scientific ideas rarely start from a blank page. They inherit mechanisms, repair known limitations, and recombine pieces of earlier work, much like biol

model-releasesarxiv-cs-ai
10 Jul 2026
← Previous
1…7576777879…361
Next →