AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,584 results
2 Jun 2026

OARelatedWork: A Large-Scale Dataset of Related Work Sections with Full-texts from Open Access Sources

Model ReleasesDGX agent

arXiv:2405.01930v2 Announce Type: replace Abstract: This paper introduces OARelatedWork: a dataset for related work generation from open-access sources. It is the first large-scale multi-document summ

OCC-RAG: Optimal Cognitive Core for Faithful Question Answering

ResearchDGX agent

arXiv:2606.00683v1 Announce Type: new Abstract: Recent progress in the development of language models has been defined by scale, with each generation absorbing more of the world's knowledge into its w

OctoT2I: A Self-Evolving Agentic Text-to-Image Router

AgentsDGX agent

arXiv:2606.01803v1 Announce Type: new Abstract: The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns fro

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On Wednesdays, We Ask Questions: Optimizing 'Active Listening' in Automated Legal Triage and Referral

Model ReleasesDGX agent

arXiv:2606.00272v1 Announce Type: new Abstract: The FETCH classifier generates follow-up questions to help refine the best match for the applicant's legal problem, using a low-cost ensemble of LLMs. I

OptiWorld: Optimal Control for Video World Generation under Physical Constraints

ResearchDGX agent

arXiv:2606.00499v1 Announce Type: new Abstract: Video generation models are becoming a scalable form of world models, but they mainly generate plausible motion rather than proactively control or optim

Parallel Complex Diffusion for Scalable Time Series Generation

TutorialsDGX agent

arXiv:2602.17706v2 Announce Type: replace Abstract: Diffusion models learn data distributions indirectly through denoising, making the difficulty of generative modeling closely tied to the dependency

Physics-Guided Attention in a Lightweight TCN for Efficient WiFi CSI-Based Human Activity Recognition

Model ReleasesDGX agent

arXiv:2606.01834v1 Announce Type: cross Abstract: Human Action Recognition (HAR) using WiFi Channel State Information (CSI) has gained increasing attention due to its non-contact, low-cost, and privac

Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs

ResearchDGX agent

arXiv:2606.01858v1 Announce Type: new Abstract: Users increasingly expect image generation models to quickly adapt to highly diverse and personalized requirements, such as producing images with distin

Position: Current Benchmarking Hinders Real Progress in Deep Learning for Time Series Forecasting

Local AiDGX agent

arXiv:2512.22702v2 Announce Type: replace Abstract: Deep learning models have grown popular in time series applications. However, the large quantity of newly proposed architectures and the often contr

Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation

Model ReleasesDGX agent

arXiv:2606.00629v1 Announce Type: cross Abstract: Sound design workflows frequently oscillate between time-consuming library searches and the complexity of procedural synthesis, with practitioners typ

RadioMaster: Multi-Agent System for Autonomous Radio Signal Generation

Model ReleasesDGX agent

arXiv:2606.01862v1 Announce Type: cross Abstract: Translating user intents into physical radio signals represents the critical yet notoriously tedious final step in wireless prototyping, as it require

Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs

ResearchDGX agent

arXiv:2606.01620v1 Announce Type: new Abstract: Video diffusion models have significantly advanced portrait video generation, yet their high computational demands limit their use in interactive applic

Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems

AgentsDGX agent

arXiv:2606.01351v1 Announce Type: new Abstract: The transition from single-turn models to Multi-Agent Systems (MAS) promises enhanced problem-solving capabilities, yet the centralized orchestration to

RedDebate: Safer Responses Through Multi-Agent Red Teaming Debates

SafetyDGX agent

arXiv:2506.11083v3 Announce Type: replace Abstract: We introduce RedDebate, a novel multi-agent debate framework that provides the foundation for Large Language Models (LLMs) to identify and mitigate

Ryze: Evidence-Enriched Data Synthesis from Biomedical Papers

Model ReleasesDGX agent

arXiv:2606.00902v1 Announce Type: new Abstract: General-purpose VLMs remain unreliable for biomedical research because valid answers in scientific papers depend on evidence split across figures, table

Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization

SafetyDGX agent

arXiv:2510.09330v3 Announce Type: replace Abstract: Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches pr

Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-Tuning and Can Be Mitigated by Machine Unlearning

SafetyDGX agent

arXiv:2503.11832v5 Announce Type: replace Abstract: Recent vision language models (VLMs) have made remarkable strides in generative modeling with multimodal inputs, particularly text and images. Howev

SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

SafetyDGX agent

arXiv:2602.16220v2 Announce Type: replace Abstract: Modeling multiscale patterns is crucial for long-term time series forecasting (TSF). However, redundancy and noise in time series, together with sem

The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Shortcuts in LLMs

SafetyDGX agent

arXiv:2606.00674v1 Announce Type: cross Abstract: Large Language Models (LLMs) aligned via outcome-based Reinforcement Learning (RL) frequently exhibit a critical failure mode: they achieve high perfo

Through the PRISM: Principle-Aware, Interpretable, and Multi-Scale Evaluation of Visual Designs

Model ReleasesDGX agent

arXiv:2606.00592v1 Announce Type: new Abstract: Effective visual communication stems from the harmony of multiple design principles, such as readability, contrast, alignment, overlap, and coherence, w

Towards Anytime Retrieval: A Benchmark for Anytime Person Re-Identification

Model ReleasesDGX agent

arXiv:2509.16635v2 Announce Type: replace Abstract: In real applications, person re-identification (ReID) is expected to retrieve the target person at any time, including both daytime and nighttime, r

Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval

SafetyDGX agent

arXiv:2606.01615v1 Announce Type: new Abstract: Video-language models are pivotal for tasks such as moment retrieval and highlight detection, yet they often struggle to capture the dynamic, non-linear

UniD^3: A Knowledge Graph-Enhanced RAG Framework for Drug-Disease Discovery and Reasoning

Model ReleasesDGX agent

arXiv:2606.01394v1 Announce Type: new Abstract: Systematic characterization of drug-disease relationships is essential for drug discovery and repurposing, yet is hindered by the heterogeneity and rapi

v0.30.1

Local AiDGX agent

Ollama 0.30 provides improved compatibility and performance using llama.cpp, augments the MLX engine on Apple Silicon for broader hardware support, and brings support for a wider range of models inclu

v0.30.1: llm: ignore llama-server SSE ping comments (#16443)

Model ReleasesDGX agent

Ollama v0.30.1 addresses an issue where the LLM component now ignores Server-Sent Events (SSE) ping comments from llama-server, resolving problem #16443. This fix improves the stability and reliabilit

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1.…

Model ReleasesDGX agent

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1. MSA (MiniMax Sparse Attention) is the star ⭐️. Unlike CSA/HC

Where Do Deep-Research Agents Go Wrong? Span-Level Error Localization in Agent Trajectories

Model ReleasesDGX agent

arXiv:2606.02060v1 Announce Type: new Abstract: Deep-research agents solve tasks through long trajectories of search, tool use, evidence inspection, and answer synthesis. Evaluation based on final ans

1 Jun 2026

A Lightweight Ensemble-Based Face Image Quality Assessment Method with Correlation-Aware Loss

Model ReleasesDGX agent

arXiv:2509.10114v2 Announce Type: replace Abstract: Face image quality assessment (FIQA) plays a critical role in face recognition and verification systems, especially in uncontrolled, real-world envi

A Unified and Reproducible Experimentation Framework for Speech Understanding

AgentsDGX agent

arXiv:2605.30899v1 Announce Type: cross Abstract: Speech foundation models and Speech LLMs have advanced speech understanding, yet deployment-oriented model selection is hindered by non-comparable eva

AbstainGNN: Teaching Graph Neural Networks to Abstain for Graph Classification

Model ReleasesDGX agent

arXiv:2605.30786v1 Announce Type: new Abstract: Graph classification is a core task in graph data mining with widespread real-world applications. Recent advances in graph neural networks (GNNs) have l

Biases in the Blind Spot: Detecting What LLMs Fail to Mention

SafetyDGX agent

arXiv:2602.10117v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) often provide chain-of-thought (CoT) reasoning traces that appear plausible, but may hide internal biases. We cal

Calibrated Preference Learning: The Case of Label Ranking

Model ReleasesDGX agent

arXiv:2605.30447v1 Announce Type: cross Abstract: Calibration, the alignment of predicted probabilities with true outcome frequencies, is essential for reliable decision-making. While extensively stud

CanLegalRAGBench: Evaluating Retrieval-Augmented Generation on Canadian Case Law

Model ReleasesDGX agent

arXiv:2605.30497v1 Announce Type: new Abstract: RAG-based legal assistants have been growing in popularity, but LLM hallucinations remain a key issue and potentially undermines justice. While benchmar

Chem-PerturBridge: a harmonized compendium of small molecule perturbation transcriptomic effects

ResearchDGX agent

arXiv:2605.31522v1 Announce Type: new Abstract: Large perturbation models require training data encompassing chemical, cellular, and assay diversity. Current transcriptomic resources for small-molecul

Cosmos3-Super-Image2Video running locally on a single RTX PRO 6000 96GB

Local AiDGX agent

Cosmos3-Super-Image2Video is an omnimodal world model capable of generating video from combinations of text and image inputs . The model is designed to run on workstation-grade compute like the NVIDIA

Dex2HOI: Dexterous Bimanual Two-Object Interaction Generation

Model ReleasesDGX agent

arXiv:2605.30444v1 Announce Type: new Abstract: Recent advances in 4D Human-Object Interaction (HOI) generation have enabled increasingly realistic motion synthesis, particularly for single-object man

Do covariates explain why these groups differ? The choice of reference group can reverse conclusions in the Oaxaca-Blinder decomposition

Model ReleasesDGX agent

arXiv:2603.29972v2 Announce Type: replace-cross Abstract: Scientists often want to explain why an outcome is different in two groups. For instance, differences in patient mortality rates across two ho

DTBench: A Synthetic Benchmark for Document-to-Table Extraction

Model ReleasesDGX agent

arXiv:2602.13812v3 Announce Type: replace-cross Abstract: Document-to-table (Doc2Table) extraction derives structured tables from unstructured documents under a target schema, enabling reliable and ve

EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision

Model ReleasesDGX agent

arXiv:2605.31557v1 Announce Type: new Abstract: Continuous episodic memory is a core capability for autonomous agents operating in dynamic, real-world environments, yet current streaming video benchma

EHRBench: An Automated and Reliable EHR-based Benchmark for Clinical Decision Making with LLMs

Model ReleasesDGX agent

arXiv:2605.30637v1 Announce Type: new Abstract: Clinical decision-making (CDM) is central to real-world clinical workflows, where clinicians infer diagnoses, select treatments, or anticipate future he

From Out-of-Distribution Detection to Hallucination Detection: A Geometric View

SafetyDGX agent

arXiv:2602.07253v2 Announce Type: replace Abstract: Detecting hallucinations in large language models is a critical open problem with significant implications for safety and reliability. While existin

Gap-K%: Measuring Top-1 Prediction Gap for Detecting Pretraining Data

Local AiDGX agent

arXiv:2601.19936v2 Announce Type: replace-cross Abstract: The opacity of massive pretraining corpora in Large Language Models (LLMs) raises significant privacy and copyright concerns, making pretraini

GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration

ApplicationsDGX agent

arXiv:2605.31039v1 Announce Type: new Abstract: Real-world image restoration (IR) is bottlenecked by the scarcity of high-quality paired training data. Synthetic datasets are abundant but often fail t

Go-UT-Bench: A Fine-Tuning Dataset for LLM-Based Unit Test Generation in Go

Model ReleasesDGX agent

arXiv:2511.10868v2 Announce Type: replace Abstract: Training data imbalance poses a major challenge for code LLMs. Most available data heavily over represents raw opensource code while underrepresenti

Interpretability Without Tradeoffs: Disentangling Polysemanticity At Equal Predictive Performance

TutorialsDGX agent

arXiv:2605.31304v1 Announce Type: cross Abstract: Deep neural networks (DNNs) are widely used, but interpreting what they actually learn remains difficult. A major obstacle is that individual neurons

LinTree: Improving LLM Reasoning with Explicitly Structured Search Histories

Local AiDGX agent

arXiv:2605.31492v1 Announce Type: new Abstract: Large language models (LLMs) often solve reasoning problems by generating intermediate traces that explore and revise partial solutions. From a search p

LLMs Without Deep Neural Networks: New Architecture, Benefits and Case Study

Model ReleasesDGX agent

arXiv:2605.30385v1 Announce Type: cross Abstract: The purpose of this article is to provide validation to my deep neural network alternative in the context of LLMs. Very recently, there has been a sig

Local AI News You Missed - May 2026

Local AiDGX agent

This Reddit post from r/StableDiffusion likely curates May 2026 AI news highlights, including releases like Stable Audio 3.0, a model family for artistic experimentation with open-weight models. The p

LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

Model ReleasesDGX agent

arXiv:2605.30434v1 Announce Type: cross Abstract: Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability

MAVEN: Improving Generalization in Agentic Tool Calling

Model ReleasesDGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

MiniMax M3 imminent. Will be doing deep testing with it on my own coding agent and harness. Review coming soon.

AgentsDGX agent

MiniMax M3, an upcoming AI model, is expected to be released soon and will undergo comprehensive testing within a custom coding agent framework. A detailed technical review of the model's performance

MiniMax M3 launched!

Local AiDGX agent

MiniMax M3 launched on June 1, 2026 as the first open-weights model to combine frontier coding, a 1-million-token context window, and native multimodality. The model achieves top-tier performance on c

On-Device Generative AI for GDPR-Compliant Visual Monitoring: Natural Language Alerts from Local Object Detection

Model ReleasesDGX agent

arXiv:2605.30544v1 Announce Type: new Abstract: Visual monitoring systems that rely on cloud-based AI inference expose raw image data to external services, creating fundamental tensions with the data-

OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning

ApplicationsDGX agent

arXiv:2605.30736v1 Announce Type: cross Abstract: The rapid development of large language models, each with distinct capabilities and inference costs, raises a practical deployment question: given an

PictSure: Pretraining Embeddings Matters for In-Context Learning Image Classifiers

AgentsDGX agent

arXiv:2506.14842v2 Announce Type: replace-cross Abstract: Building image classification models remains cumbersome in data-scarce domains, where collecting large labeled datasets is impractical. In-con

PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection

ApplicationsDGX agent

arXiv:2502.12119v4 Announce Type: replace-cross Abstract: Visual instruction tuning adapts pre-trained Multimodal Large Language Models (MLLMs) to follow human instructions for real-world applications

Recognizing Co-Speech Gestures in-the-Wild

Model ReleasesDGX agent

arXiv:2605.31589v1 Announce Type: new Abstract: While humans naturally gesture during speech, only a sparse subset of these movements are visually depictive and semantically linked to specific spoken

Rectified flow-based prediction of post-treatment brain MRI from pre-radiotherapy priors for patients with glioma

ResearchDGX agent

arXiv:2603.08385v2 Announce Type: replace-cross Abstract: Brain tumors result in 20 years of lost life on average. Standard therapies induce complex structural changes in the brain that are monitored

Scaling Conversational Hungarian ASR: The BEA-Dialogue+ Corpus

Model ReleasesDGX agent

arXiv:2605.31469v1 Announce Type: cross Abstract: Conversational automatic speech recognition in Hungarian is constrained by the limited amount of publicly available dialogue-style training data. The

Scaling Higher-Order Graph Learning with Maximal Clique Complexes

ResearchDGX agent

arXiv:2605.31373v1 Announce Type: cross Abstract: Graph neural networks (GNNs) are limited to modeling pairwise interactions, while higher-order models based on cell complexes achieve greater expressi

← Previous
1…465466467468469…1060
Next →