AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Omni2Sound: Towards Unified Video-Text-to-Audio Generation

DGX agent

arXiv:2601.02731v3 Announce Type: replace-cross Abstract: Training a unified model integrating video-to-audio (V2A), text-to-audio (T2A), and joint video-text-to-audio (VT2A) generation offers signifi

model-releasesarxiv-cs-cv
30 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Option-Order Randomisation Reveals a Distributional Position Attractor in Prompted Sandbagging

DGX agent

arXiv:2604.26206v1 Announce Type: cross Abstract: A predecessor pilot (Cacioli, 2026) found that Llama-3-8B implements prompted sandbagging as positional collapse rather than answer avoidance. However

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Parameterized Quantum Circuits as Feature Maps: Representation Quality and Readout Effects in Multispectral Land-Cover Classification

DGX agent

arXiv:2604.26675v1 Announce Type: cross Abstract: We investigate variational quantum classifiers (VQCs) for land-cover classification from multispectral satellite imagery, adopting a feature-map persp

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

PATCH: Learnable Tile-level Hybrid Sparsity for LLMs

DGX agent

arXiv:2509.23410v4 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver impressive performance but incur prohibitive memory and compute costs at deployment. Model pruning is an

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Perception Test 2025: Challenge Summary and a Unified VQA Extension

DGX agent

arXiv:2601.06287v2 Announce Type: replace Abstract: The Third Perception Test challenge was organised as a full-day workshop alongside the IEEE/CVF International Conference on Computer Vision (ICCV) 2

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation

DGX agent

arXiv:2604.26561v1 Announce Type: cross Abstract: Multi-agent deliberation systems using large language models (LLMs) are increasingly proposed for policy simulation, yet they suffer from artificial c

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models

DGX agent

arXiv:2604.26508v1 Announce Type: cross Abstract: Deploying Vision-Language Models (VLMs) on edge devices remains challenging due to their substantial computational and memory demands, which exceed th

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

QERNEL: a Scalable Large Electron Model

DGX agent

arXiv:2604.26018v1 Announce Type: cross Abstract: We introduce QERNEL, a foundational neural wavefunction that variationally solves families of parameterized many-electron Hamiltonians and captures th

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Quantum Feature Selection with Higher-Order Binary Optimization on Trapped-Ion Hardware

DGX agent

arXiv:2604.26834v1 Announce Type: cross Abstract: We present a quantum feature-selection framework based on a higher-order unconstrained binary optimization (HUBO) formulation that explicitly incorpor

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing

DGX agent

arXiv:2604.26435v1 Announce Type: cross Abstract: The rapid advancement of object detection architectures has positioned single stage detectors as the dominant solution for real-time visual perception

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments

DGX agent

arXiv:2604.26067v1 Announce Type: new Abstract: We present RADIO-ViPE (Reduce All Domains Into One -- Video Pose Engine), an online semantic SLAM system that enables geometry-aware open-vocabulary gro

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

RaMP: Runtime-Aware Megakernel Polymorphism for Mixture-of-Experts

DGX agent

arXiv:2604.26039v1 Announce Type: cross Abstract: The optimal kernel configuration for Mixture-of-Experts (MoE) inference depends on both batch size and the expert routing distribution, yet production

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Random Cloud: Finding Minimal Neural Architectures Without Training

DGX agent

arXiv:2604.26830v1 Announce Type: cross Abstract: I propose the Random Cloud method, a training-free approach to neural architecture search that discovers minimal feedforward network topologies throug

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Reasoning Gets Harder for LLMs Inside A Dialogue

DGX agent

arXiv:2603.20133v2 Announce Type: replace Abstract: Large Language Models (LLMs) achieve strong performance on many reasoning benchmarks, yet these evaluations typically focus on isolated tasks that d

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization

DGX agent

arXiv:2602.15983v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that execut

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Retrieval-Augmented LLMs for Evidence Localization in Clinical Trial Recruitment from Longitudinal EHR Narratives

DGX agent

arXiv:2604.05190v2 Announce Type: replace-cross Abstract: Screening patients for enrollment is a well-known, labor-intensive bottleneck that leads to under-enrollment and, ultimately, trial failures.

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

reward-lens: A Mechanistic Interpretability Library for Reward Models

DGX agent

arXiv:2604.26130v1 Announce Type: cross Abstract: Every RLHF-trained language model is shaped by a reward model, yet the mechanistic interpretability toolkit -- logit lens, direct logit attribution, a

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment

DGX agent

arXiv:2601.04389v2 Announce Type: replace-cross Abstract: Current safety evaluations of large language models (LLMs) create a dangerous illusion of universal protection by aggregating harms under gene

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

SciMDR: Advancing Scientific Multimodal Document Reasoning

DGX agent

arXiv:2603.12249v2 Announce Type: replace-cross Abstract: Constructing scientific multimodal document reasoning datasets for foundation model training involves an inherent trade-off among scale, faith

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

SEAL: Semantic-aware Single-image Sticker Personalization with a Large-scale Sticker-tag Dataset

DGX agent

arXiv:2604.26883v1 Announce Type: new Abstract: Synthesizing a target concept from a single reference image is challenging in diffusion-based personalized text-to-image generation, particularly for st

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Self-Jailbreaking: Language Models Can Reason Themselves Out of Safety Alignment After Benign Reasoning Training

DGX agent

arXiv:2510.20956v2 Announce Type: replace-cross Abstract: We discover a novel and surprising phenomenon of unintentional misalignment in reasoning language models (RLMs), which we call self-jailbreaki

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Shorthand for Thought: Compressing LLM Reasoning via Entropy-Guided Supertokens

DGX agent

arXiv:2604.26355v1 Announce Type: new Abstract: Reasoning in Large Language Models incurs significant inference-time compute, yet the token-level information structure of reasoning traces remains unde

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

SongBench: A Fine-Grained Multi-Aspect Benchmark for Song Quality Assessment

DGX agent

arXiv:2604.25937v1 Announce Type: cross Abstract: Recent advancements in Text-to-Song generation have enabled realistic musical content production, yet existing evaluation benchmarks lack the professi

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

StarDrinks: An English and Korean Test Set for SLU Evaluation in a Drink Ordering Scenario

DGX agent

arXiv:2604.26500v1 Announce Type: new Abstract: LLMs and speech assistants are increasingly used for task-oriented interactions, yet their evaluation often relies on controlled scenarios that fail to

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading

DGX agent

arXiv:2604.26614v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on general multimodal tasks, yet they remain brittle on dial-based measuremen

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

StratMem-Bench: Evaluating Strategic Memory Use in Virtual Character Conversation Beyond Factual Recall

DGX agent

arXiv:2604.26243v1 Announce Type: cross Abstract: Achieving realistic human-like conversation for virtual characters requires not only a simple memorization and recall of past events, but also the str

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Stress Testing Factual Consistency Metrics for Long-Document Summarization

DGX agent

arXiv:2511.07689v2 Announce Type: replace-cross Abstract: Evaluating the factual consistency of abstractive text summarization remains a significant challenge, particularly for long documents, where c

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Structural Generalization on SLOG without Hand-Written Rules

DGX agent

arXiv:2604.26157v1 Announce Type: cross Abstract: Structural generalization in semantic parsing requires systems to apply learned compositional rules to novel structural combinations. Existing approac

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

SWE-Edit: Rethinking Code Editing for Efficient SWE-Agent

DGX agent

arXiv:2604.26102v1 Announce Type: cross Abstract: Large language model agents have achieved remarkable progress on software engineering tasks, yet current approaches suffer from a fundamental context

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection

DGX agent

arXiv:2604.26772v1 Announce Type: new Abstract: Recent methods demonstrate that large-scale pretrained models, such as CLIP vision transformers, effectively detect AI-generated images (AIGIs) from uns

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

The Unseen Adversaries: Robust and Generalized Defense Against Adversarial Patches

DGX agent

arXiv:2604.26317v1 Announce Type: new Abstract: The vulnerabilities of deep neural networks against singularities have raised serious concerns regarding their deployment in the physical world. One of

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Theory-Grounded Evaluation Exposes the Authorship Gap in LLM Personalization

DGX agent

arXiv:2604.26460v1 Announce Type: new Abstract: Stylistic personalization - making LLMs write in a specific individual's style, rather than merely adapting to task preferences - lacks evaluation groun

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Thinking with Drafting: Optical Decompression via Logical Reconstruction

DGX agent

arXiv:2602.11731v2 Announce Type: replace Abstract: Existing multimodal large language models have achieved high-fidelity visual perception and exploratory visual generation. However, a precision para

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

TildeOpen LLM: Leveraging Curriculum Learning to Achieve Equitable Language Representation

DGX agent

arXiv:2603.08182v2 Announce Type: replace-cross Abstract: Large language models often underperform in many European languages due to the dominance of English and a few high-resource languages in train

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time Blindness: Why Video-Language Models Can't See What Humans Can?

DGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time series classification with random convolution kernels: pooling operators and input representations matter

DGX agent

arXiv:2409.01115v5 Announce Type: replace Abstract: This article presents a new approach based on MiniRocket, called SelF-Rocket, for fast time series classification (TSC). Unlike existing approaches

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

TinyR1-32B-Preview: Boosting Accuracy with Branch-Merge Distillation

DGX agent

arXiv:2503.04872v3 Announce Type: replace-cross Abstract: The challenge of reducing the size of Large Language Models (LLMs) while maintaining their performance has gained significant attention. Howev

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Training-Free Adaptation of New-Generation LLMs using Legacy Clinical Models

DGX agent

arXiv:2601.03423v3 Announce Type: replace-cross Abstract: Adapting language models to the clinical domain through continued pretraining and instruction tuning requires costly retraining for each new m

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Training-Free Loosely Speculative Decoding: Accepting Semantically Correct Drafts Beyond Exact Match

DGX agent

arXiv:2511.22972v3 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse tasks but suffer from high inference latency due to their autoregressive gene

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

DGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models

DGX agent

arXiv:2505.22897v2 Announce Type: replace Abstract: While bias in large language models (LLMs) is well-studied, similar concerns in vision-language models (VLMs) have received comparatively less atten

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness

DGX agent

arXiv:2603.07080v3 Announce Type: replace-cross Abstract: Vision-and-Language Navigation (VLN) increasingly relies on large vision-language models, but their inference cost conflicts with real-time de

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

VulStyle: A Multi-Modal Pre-Training for Code Stylometry-Augmented Vulnerability Detection

DGX agent

arXiv:2604.26313v1 Announce Type: cross Abstract: We present VulStyle, a multi-modal software vulnerability detection model that jointly encodes function-level source code, non-terminal Abstract Synta

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models

DGX agent

arXiv:2510.14438v2 Announce Type: replace Abstract: The hallmark of Deep Research agents lies in compositional reasoning, the capacity to aggregate distributed, heterogeneous information into coherent

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

DGX agent

arXiv:2604.26649v1 Announce Type: cross Abstract: Large reasoning models such as DeepSeek-R1 and OpenAI o1 generate extended chains of thought spanning thousands of tokens, yet their integration with

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

A Comparative Analysis on the Performance of Upper Confidence Bound Algorithms in Adaptive Deep Neural Networks

DGX agent

arXiv:2604.24810v1 Announce Type: new Abstract: Edge computing environments impose strict constraints on energy consumption and latency, making the deployment of deep neural networks a significant cha

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

A Comparative Study in Surgical AI: Datasets, Foundation Models, and Barriers to Med-AGI

DGX agent

arXiv:2603.27341v2 Announce Type: replace-cross Abstract: Recent Artificial Intelligence (AI) models have matched or exceeded human experts in several benchmarks of biomedical task performance, but su

model-releasesarxiv-cs-cv
29 Apr 2026
← Previous
1…290291292293294…361
Next →