AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,036 results
Model Releases

RAS: Measuring LLM Safety Through Refusal Alignment

DGX agent

arXiv:2606.25750v1 Announce Type: cross Abstract: Safety evaluation of large language models (LLMs) is commonly performed by querying models with unsafe or jailbreak prompts and judging whether their

model-releasesarxiv-cs-cl
25 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Riazi-8B: An Urdu Large Language Model for Mathematical Reasoning

DGX agent

arXiv:2606.25568v1 Announce Type: new Abstract: Recent LLMs demonstrate strong mathematical reasoning capabilities, but existing gains rely heavily on English-centric training resources and benchmarks

researcharxiv-cs-cl
25 Jun 2026
Tools

Which tokens does a hybrid model predict better?

DGX agent

This article from Allen AI discusses how hybrid models that combine different prediction approaches perform across various token types, likely comparing their effectiveness on common tokens versus rar

toolshugging-face
25 Jun 2026
Safety

A Robust Model-Based Approach for Continuous-Time Policy Evaluation with Unknown Levy Process Dynamics

DGX agent

arXiv:2504.01482v3 Announce Type: replace-cross Abstract: This paper develops a model-based framework for continuous-time policy evaluation (CTPE) in reinforcement learning, incorporating both Brownia

safetyarxiv-cs-lg
24 Jun 2026
Research

Catastrophic Compositional Generation: Why Vanilla Diffusion Models Fail to Extrapolate

DGX agent

arXiv:2606.23920v1 Announce Type: cross Abstract: The task of compositional generation involves using a conditional generative model, trained only on a subset of the possible conditions, to produce sa

researcharxiv-cs-ai
24 Jun 2026
Applications

Ensemble Learning for Large Language Models in Text and Code Generation: A Survey

DGX agent

arXiv:2503.13505v3 Announce Type: replace-cross Abstract: Generative Pretrained Transformers (GPTs) are foundational Large Language Models (LLMs) for text generation. However, individual LLMs often pr

applicationsarxiv-cs-ai
24 Jun 2026
Local Ai

ForensicsTok: Forensics-Guided Tokenized Modeling for Image Tampering Localization

DGX agent

arXiv:2606.24538v1 Announce Type: new Abstract: Multi-modal Large Language Models (MLLMs) offer powerful reasoning for forensic tasks, yet existing approaches utilizing exogenous segmentation decoders

local-aiarxiv-cs-cv
24 Jun 2026
Safety

From 'Aha Moments' to Controllable Thinking: Toward Meta-Cognitive Reasoning in Large Reasoning Models via Decoupled Reasoning and Control

DGX agent

arXiv:2508.04460v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) can exhibit step-by-step reasoning, reflection, and backtracking, but these behaviors are often unregulated, leading t

safetyarxiv-cs-ai
24 Jun 2026
Applications

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A lea…

DGX agent

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A leaderboard winner can still lose on your codebase. Training cl

applicationsfireworks-ai--x
24 Jun 2026
Model Releases

Introducing Claude for Music. You can now create songs from Claude Code, Hermes, Codex, or any agent you’re using. SOTA music model @MiniMax…

DGX agent

I cannot provide an accurate summary for this entry. The URL and source attribution appear inconsistent (title credits Yohei Nakajima but URL references a different user), and the post references prod

model-releasesyohei-nakajima--x
24 Jun 2026
Safety

On the Smallness of the Large Language Models Scaling Exponents

DGX agent

arXiv:2606.24504v1 Announce Type: new Abstract: We discuss reasons why the scaling exponents of current Large Language Models (LLMs) applications are indicating an unsustainable regime in terms of ene

safetyarxiv-cs-ai
24 Jun 2026
Research

Separating Oblivious and Adaptive Models of Variable Selection

DGX agent

arXiv:2602.16568v2 Announce Type: replace-cross Abstract: Sparse recovery is among the most well-studied problems in learning theory and high-dimensional statistics. In this work, we investigate the s

researcharxiv-cs-lg
24 Jun 2026
Safety

Spectral Evolution-Guided Token Pruning in Multimodal Large Language Models

DGX agent

arXiv:2606.24165v1 Announce Type: new Abstract: Reducing visual token redundancy is critical for accelerating Multimodal Large Language Models (MLLMs) without degrading cross-modal reasoning performan

safetyarxiv-cs-cv
24 Jun 2026
Research

Stabilizing Physics-Informed Consistency Models via Structure-Preserving Training

DGX agent

arXiv:2602.09303v2 Announce Type: replace Abstract: We propose a physics-informed consistency modeling framework for solving partial differential equations (PDEs) via fast, few-step generative inferen

researcharxiv-cs-lg
24 Jun 2026
Agents

Unlimited OCR is a great model on table parsing and understanding proper reading order. However it does struggle a little on semantic format…

DGX agent

Unlimited OCR is a great model on table parsing and understanding proper reading order. However it does struggle a little on semantic formatting, charts (it does decent at bounding boxes). Attaching t

agentsjerry-liu--x
24 Jun 2026
Safety

UOL@IDEM at BEA 2026 Shared Task 1: Neural Fusion and Feature-Rich Modeling for L1-Aware Vocabulary Difficulty Prediction

DGX agent

arXiv:2606.24501v1 Announce Type: new Abstract: This paper describes UOL@IDEM's closed-track submission to the BEA 2026 shared task on L1-aware vocabulary difficulty prediction. We model the task as r

safetyarxiv-cs-cl
24 Jun 2026
Research

What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning

DGX agent

arXiv:2506.00869v3 Announce Type: replace Abstract: Despite the impressive performance of vision-language models (VLMs) on downstream tasks, their ability to understand and reason about causal relatio

researcharxiv-cs-cl
24 Jun 2026
Research

A Comprehensive Study on Visual Token Redundancy for Discrete Diffusion-based Multimodal Large Language Models

DGX agent

arXiv:2511.15098v2 Announce Type: replace Abstract: Discrete diffusion-based multimodal large language models (dMLLMs) have emerged as a promising alternative to autoregressive MLLMs thanks to their a

researcharxiv-cs-cv
23 Jun 2026
Research

An Effective Strategy for Modeling Score Ordinality and Non-uniform Intervals in Automated Speaking Assessment

DGX agent

arXiv:2509.03372v3 Announce Type: replace-cross Abstract: A recent line of research on automated speaking assessment (ASA) has benefited from self-supervised learning (SSL) representations, which capt

researcharxiv-cs-lg
23 Jun 2026
Tutorials

Bayesian Model Averaging under Predictor Redundancy via Density-Ratio Posterior Compression

DGX agent

arXiv:2606.21080v1 Announce Type: cross Abstract: Bayesian model averaging in support-indexed regression induces a posterior distribution over active predictor supports. Under predictor redundancy, po

tutorialsarxiv-cs-lg
23 Jun 2026
Industry

ByteDance unveils Seedance 2.5 in Beijing, saying the AI video model can generate 30-second clips from up to 50 reference materials, up from 12 for Seedance 2.0 (Juro Osawa/The Information)

DGX agent

Juro Osawa / The Information: ByteDance unveils Seedance 2.5 in Beijing, saying the AI video model can generate 30-second clips from up to 50 reference materials, up from 12 for Seedance 2.0 — ByteDan

industrytechmeme
23 Jun 2026
Research

Compression and Retrieval: Implicit Memory Retrieval for Video World Models

DGX agent

arXiv:2606.23105v1 Announce Type: new Abstract: Video world models hold promise for simulating interactive environments, yet maintaining consistent long-term memory across complex camera trajectories

researcharxiv-cs-cv
23 Jun 2026
Research

Cultural Counterfactuals: Evaluating Cultural Biases in Large Vision-Language Models with Counterfactual Examples

DGX agent

arXiv:2603.02370v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have grown increasingly powerful in recent years, but can also exhibit harmful biases. Prior studies investigat

researcharxiv-cs-cv
23 Jun 2026
Safety

Data-Driven Image Registration and Deformation Modeling for Image-Guided Neurosurgery: A Systematic Review

DGX agent

arXiv:2602.10155v2 Announce Type: replace-cross Abstract: Accurate compensation of brain deformation is critical for reliable image-guided neurosurgery. Surgical manipulation and tumor resection induc

safetyarxiv-cs-cv
23 Jun 2026
Research

Diffusion Models Adapt to Low-Dimensional Structure Under Flexible Coefficient Choices

DGX agent

arXiv:2606.23627v1 Announce Type: cross Abstract: Diffusion models are known to exploit unknown low-dimensional structure to accelerate sampling. However, existing convergence theory under low-dimensi

researcharxiv-cs-lg
23 Jun 2026
Safety

Extraction and Analysis of Multimodal Concepts in Vision Language Models through Sparse Autoencoders

DGX agent

arXiv:2606.21197v1 Announce Type: new Abstract: Vision Language Models (VLMs) have demonstrated impressive performance in tasks requiring joint understanding of images and text, such as image captioni

safetyarxiv-cs-cv
23 Jun 2026
Research

Fine-grained Human Motion Understanding with Language Models

DGX agent

arXiv:2606.20888v1 Announce Type: new Abstract: In this work, we propose methodname, an LLM-based model for fine-grained human motion understanding that represents motion as a sequence of skeletal pos

researcharxiv-cs-cv
23 Jun 2026
Applications

Flatness Preserves Instruction Following in Vision-Language-Action Models

DGX agent

arXiv:2606.23641v1 Announce Type: new Abstract: Vision-language-action (VLA) models have the potential for open-world generalization by leveraging pretrained vision-language representations, yet downs

applicationsarxiv-cs-ro
23 Jun 2026
Model Releases

Hierarchical Sparse Circuit Extraction from Billion-Parameter Language Models through Scalable Attribution Graph Decomposition

DGX agent

arXiv:2601.12879v2 Announce Type: replace Abstract: Extracting sparse circuits from billion-parameter transformers is constrained by O(2^n) search cost and pervasive feature reuse across co-active pat

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Interpretable Probabilistic Medical Image Segmentation via Gaussian Process with Explicit Modelling of Annotation Bias and Variability

DGX agent

arXiv:2606.23177v1 Announce Type: new Abstract: Deep learning-based medical image segmentation models are trained using annotations that exhibit systematic bias and variability across raters. While pr

safetyarxiv-cs-cv
23 Jun 2026
Safety

MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning

DGX agent

arXiv:2606.20641v1 Announce Type: cross Abstract: Multi-modal Large Language Models (MLLMs) have demonstrated remarkable capabilities in semantic understanding and common sense reasoning, making them

safetyarxiv-cs-lg
23 Jun 2026
Local Ai

MambaADv2: Evolving Duality-enhanced State Space Model for Unsupervised Anomaly Detection

DGX agent

arXiv:2606.23126v1 Announce Type: new Abstract: While recent advancements in anomaly detection have demonstrated the efficacy of CNN- and Transformer-based approaches, these architectures face inheren

local-aiarxiv-cs-cv
23 Jun 2026
Safety

Measuring Model-Induced Discrimination via Efficient Fairness Approximation

DGX agent

arXiv:2405.09251v2 Announce Type: replace Abstract: Providing various machine learning (ML) applications in the real world, concerns about discrimination hidden in ML models are growing, particularly

safetyarxiv-cs-lg
23 Jun 2026
Safety

Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

DGX agent

arXiv:2505.12462v3 Announce Type: replace Abstract: Robust reinforcement learning (RL) under the average-reward criterion is essential for long-term decision-making, particularly when the environment

safetyarxiv-cs-lg
23 Jun 2026
Local Ai

Multi-AUV Marine Life Tracking with Single Hydrophone Payloads via a Hidden Markov Model Equipped Particle Filter

DGX agent

arXiv:2606.22335v1 Announce Type: new Abstract: Researchers tag and track marine animals to study migration patterns, human impacts on behavior, and behavioral shifts due to climate change. Accurate d

local-aiarxiv-cs-ro
23 Jun 2026
Research

PeLAP-A: Adaptive Latent Pruning for Lightweight Latent Diffusion Models

DGX agent

arXiv:2606.23086v1 Announce Type: new Abstract: Latent diffusion models achieve strong generative performance by operating in a compressed latent space produced by a variational autoencoder (VAE). How

researcharxiv-cs-lg
23 Jun 2026
Safety

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics

DGX agent

arXiv:2606.22806v1 Announce Type: new Abstract: Synthesizing realistic Human-Object Interactions (HOI) is critical for creating embodied avatars and functional virtual environments. However, current d

safetyarxiv-cs-cv
23 Jun 2026
Research

Protocol-Aware Tokenization and Architecture Co-Design for Wireless Packet Foundation Models

DGX agent

arXiv:2606.20587v1 Announce Type: cross Abstract: What matters more for building foundation models for wireless packet traces: the tokenizer or the architecture or both? To answer this question, we bu

researcharxiv-cs-lg
23 Jun 2026
Local Ai

Read the technical paper on Krea 2 https://www.krea.ai/blog/krea-2-technical-report Download the model weights https://github.com/krea-ai/kr…

DGX agent

Krea 2 is a technical advancement in AI image generation with newly released model weights available for download on GitHub. The technical report details the improvements and capabilities of this vers

local-aicomfyui--x
23 Jun 2026
Safety

RelightAnyone: A Generalized Relightable 3D Gaussian Head Model

DGX agent

arXiv:2601.03357v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has become a standard approach to reconstruct and render photorealistic 3D head avatars. A major challenge is to religh

safetyarxiv-cs-cv
23 Jun 2026
Research

Revisiting OmniAnomaly for Anomaly Detection: performance metrics and comparison with PCA-based models

DGX agent

arXiv:2603.18985v2 Announce Type: replace-cross Abstract: Deep learning models have become the dominant approach for multivariate time series anomaly detection (MTSAD), often reporting substantial per

researcharxiv-cs-lg
23 Jun 2026
Tutorials

Scalable Training of Spatially Grounded 2D Vision-Language Models for Radiology

DGX agent

arXiv:2606.20477v2 Announce Type: replace Abstract: We study how to train visually grounded vision-language models (VLMs) for radiology without manual spatial annotations. We introduce RefRad2D, a lar

tutorialsarxiv-cs-cv
23 Jun 2026
Safety

Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement (New York Times)

DGX agent

New York Times: Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement — Federal officials are urg

safetytechmeme
23 Jun 2026
Applications

Token Factory: Efficiently Integrating Diverse Signals into Large Recommendation Models

DGX agent

arXiv:2606.19635v2 Announce Type: replace-cross Abstract: Large Recommendation Models (LRMs) have demonstrated promising capabilities in industry-scale recommendation tasks. However, holistically inte

applicationsarxiv-cs-lg
23 Jun 2026
Research

TooBad: Backdoor Diffusion Models with Ultra-Low Poison Rate and Imperceptible Trigger

DGX agent

arXiv:2606.23362v1 Announce Type: cross Abstract: Diffusion models (DMs), despite their impressive capabilities across a wide range of generative tasks, have been shown to be vulnerable to backdoor at

researcharxiv-cs-cv
23 Jun 2026
Safety

Training-Free Semantic Correction for Autoregressive Visual Models

DGX agent

arXiv:2606.22550v1 Announce Type: new Abstract: Autoregressive visual models (AVMs) based on next-scale prediction have emerged as a prominent paradigm for image and video synthesis. However, decompos

safetyarxiv-cs-cv
23 Jun 2026
Local Ai

Vesta: A Generalist Embodied Reasoning Model

DGX agent

arXiv:2606.20905v1 Announce Type: new Abstract: Robots operating in open-world environments must seamlessly integrate localization, spatial reasoning, navigation, and long-horizon planning. While spec

local-aiarxiv-cs-ro
23 Jun 2026
Research

When Does a Video-Language Model Stop Watching? Reward Strength Controls the Formation and Reversal of Visual Shortcuts in Multimodal RLVR

DGX agent

arXiv:2606.22043v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is increasingly applied to large vision-language models (LVLMs), yet outcome-only optimization c

researcharxiv-cs-cv
23 Jun 2026
← Previous
1…223224225226227…1272
Next →