AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
29 Apr 2026

Sharp Capacity Scaling of Spectral Optimizers in Learning Associative Memory

ResearchDGX agent

arXiv:2603.26554v2 Announce Type: replace Abstract: Spectral optimizers such as Muon have recently shown strong empirical performance in large-scale language model training, but the source and extent

Subjective Portrait Region Cropping in Landscape Videos with Temporal Annotation Smoothing

Model ReleasesDGX agent

arXiv:2604.24947v1 Announce Type: new Abstract: With the rise of mobile video consumption on diverse handheld display resolutions and orientation modes, altering videos to aspect ratios poses challeng

Use Codex to analyze a data export, flag what changed, and help draft the readout.

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

This post likely demonstrates how to use OpenAI's Codex model to programmatically analyze exported data, identify changes or anomalies between versions, and automatically generate summary reports or r

VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation

Model ReleasesDGX agent

arXiv:2604.25235v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used as automated judges for multimodal systems, yet their scores provide no indication of reliability.

Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution

SafetyDGX agent

arXiv:2604.03472v2 Announce Type: replace Abstract: Co-evolutionary self-play, where one language model generates problems and another solves them, promises autonomous curriculum learning without huma

We’re excited to introduce KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI, accepted at #I…

Model ReleasesDGX agent

We’re excited to introduce KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI, accepted at #ICASSP2026! 🐢 Blog https://pub.sakana.ai/kame/ Paper https://

28 Apr 2026

A Benchmark Suite of Reddit-Derived Datasets for Mental Health Detection

Model ReleasesDGX agent

arXiv:2604.23458v1 Announce Type: new Abstract: The growing availability of online support groups has opened up new windows to study mental health through natural language processing (NLP). However, i

Agentic AI platforms for autonomous training and rule induction of human-human and virus-human protein-protein interactions

AgentsDGX agent

arXiv:2604.23924v1 Announce Type: new Abstract: We instruct an AI agent to construct two separate agentic AI platforms: one for autonomous training of predictive ML models for human-human and virus-hu

AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents

AgentsDGX agent

arXiv:2604.24657v1 Announce Type: cross Abstract: Autonomous AI agents extend large language models into full runtime systems that load skills, ingest external content, maintain memory, plan multi-ste

AI Security Beyond Core Domains: Resume Screening as a Case Study of Adversarial Vulnerabilities in Specialized LLM Applications

Model ReleasesDGX agent

arXiv:2512.20164v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at text comprehension and generation, making them ideal for automated tasks like code review and content mo

AlphaFold's Bayesian Roots in Probability Kinematics

ResearchDGX agent

arXiv:2505.19763v3 Announce Type: replace Abstract: The seminal breakthrough of AlphaFold in protein structure prediction relied on a learned potential energy function parameterized by deep models, in

An Aircraft Upset Recovery System with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.24355v1 Announce Type: new Abstract: This article explores the progress made in the creation of a pilot activated recovery system (PARS) for advanced jet trainers that utilizes artificial i

An Analysis of Active Learning Algorithms using Real-World Crowd-sourced Text Annotations

Model ReleasesDGX agent

arXiv:2604.23290v1 Announce Type: cross Abstract: Active learning algorithms automatically identify the most informative samples from large amounts of unlabeled data and tremendously reduce human anno

ATTN-FIQA: Interpretable Attention-based Face Image Quality Assessment with Vision Transformers

Model ReleasesDGX agent

arXiv:2604.22841v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) aims to assess the recognition utility of face samples and is essential for reliable face recognition (FR) systems.

Audio2Tool: Bridging Spoken Language Understanding and Function Calling

Model ReleasesDGX agent

arXiv:2604.22821v1 Announce Type: cross Abstract: Voice assistants increasingly rely on Speech Language Models (SpeechLMs) to interpret spoken queries and execute complex tasks, yet existing benchmark

Benchmarking Emergent Coordination in Large-Scale LLM Populations: An Evaluation Framework on the MoltBook Archive

Model ReleasesDGX agent

arXiv:2603.03555v2 Announce Type: replace-cross Abstract: As multi-agent Large Language Model (LLM) systems scale, evaluating their emergent coordination dynamics becomes increasingly critical. Howeve

Boosting MLLM Spatial Reasoning with Geometrically Referenced 3D Scene Representations

Model ReleasesDGX agent

arXiv:2603.08592v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) have achieved remarkable success in 2D visual understanding, their ability to reason about 3D space r

Can Current Agents Close the Discovery-to-Application Gap? A Case Study in Minecraft

Model ReleasesDGX agent

arXiv:2604.24697v1 Announce Type: new Abstract: Discovering causal regularities and applying them to build functional systems--the discovery-to-application loop--is a hallmark of general intelligence,

Can LLMs Act as Historians? Evaluating Historical Research Capabilities of LLMs via the Chinese Imperial Examination

Model ReleasesDGX agent

arXiv:2604.24690v1 Announce Type: new Abstract: While Large Language Models (LLMs) have increasingly assisted in historical tasks such as text processing, their capacity for professional-level histori

Can You Make It Sound Like You? Post-Editing LLM-Generated Text for Personal Style

ResearchDGX agent

arXiv:2604.24444v1 Announce Type: new Abstract: Despite the growing use of large language models (LLMs) for writing tasks, users may hesitate to rely on LLMs when personal style is important. Post-edi

Cataract-LMM Large-Scale Multi-Source Multi-Task Benchmark for Deep Learning in Surgical Video Analysis

Model ReleasesDGX agent

arXiv:2510.16371v2 Announce Type: replace-cross Abstract: The development of computer-assisted surgery systems relies on large-scale, annotated datasets. Existing cataract surgery resources lack the d

CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

Model ReleasesDGX agent

arXiv:2602.23452v2 Announce Type: replace Abstract: Scientific research relies on accurate citation for attribution and integrity, yet large language models (LLMs) introduce a new risk: fabricated ref

Cloudless-Training: A Framework to Improve Efficiency of Geo-Distributed ML Training

ApplicationsDGX agent

arXiv:2303.05330v1 Announce Type: cross Abstract: Geo-distributed ML training can benefit many emerging ML scenarios (e.g., large model training, federated learning) with multi-regional cloud resource

Context management in agent harnesses: memory, files, and subagents

Model ReleasesDGX agent

A version of this article originally appeared on X. Every agent harness runs into the same limit: the context window is too small for everything the model might want to remember.... The post Context m

CorpusQA: A 10 Million Token Benchmark for Corpus-Level Analysis and Reasoning

Model ReleasesDGX agent

arXiv:2601.14952v2 Announce Type: replace-cross Abstract: While large language models now handle million-token contexts, their capacity for reasoning across entire document repositories remains largel

CrossGuard: Safeguarding MLLMs against Joint-Modal Implicit Malicious Attacks

Model ReleasesDGX agent

arXiv:2510.17687v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong reasoning and perception capabilities but are increasingly vulnerable to jailbreak att

Dr. RTL: Autonomous Agentic RTL Optimization through Tool-Grounded Self-Improvement

AgentsDGX agent

arXiv:2604.14989v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing interest in automatic RTL optimization for better performance, power, and area

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

SafetyDGX agent

arXiv:2511.13312v2 Announce Type: replace-cross Abstract: Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and

Enhanced Privacy and Communication Efficiency in Non-IID Federated Learning with Adaptive Quantization and Differential Privacy

ResearchDGX agent

arXiv:2604.23426v1 Announce Type: new Abstract: Federated learning (FL) is a distributed machine learning method where multiple devices collaboratively train a model under the management of a central

Evaluating Jailbreaking Vulnerabilities in LLMs Deployed as Assistants for Smart Grid Operations: A Benchmark Against NERC Standards

Model ReleasesDGX agent

arXiv:2604.23341v1 Announce Type: cross Abstract: The deployment of Large Language Models (LLMs) as assistants in electric grid operations promises to streamline compliance and decision-making but exp

Evaluating the Search Agent in a Parallel World

Model ReleasesDGX agent

arXiv:2603.04751v2 Announce Type: replace Abstract: Integrating web search tools has significantly extended the capability of LLMs to address open-world, real-time, and long-tail problems. However, ev

Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning

SafetyDGX agent

arXiv:2602.07605v3 Announce Type: replace-cross Abstract: Any entity in the visual world can be hierarchically grouped based on shared characteristics and mapped to fine-grained sub-categories. While

Flexible Deep Neural Networks for Partially Linear Survival Data: Estimation and Survival Inference

ResearchDGX agent

arXiv:2512.10570v2 Announce Type: replace-cross Abstract: We propose a flexible deep neural network (DNN) framework for modeling survival data within a partially linear regression structure. The appro

FlowPlace: Flow Matching for Chip Placement

ResearchDGX agent

arXiv:2604.23658v1 Announce Type: cross Abstract: Chip placement plays an important role in physical design. While generative models like diffusion models offer promising learning-based solutions, cur

Graph Memory Transformer (GMT)

Model ReleasesDGX agent

arXiv:2604.23862v1 Announce Type: cross Abstract: We investigate whether the Feed-Forward Network (FFN) sublayer in a decoder-only transformer can be replaced by an explicit learned memory graph while

GraphPlanner: Graph Memory-Augmented Agentic Routing for Multi-Agent LLMs

Model ReleasesDGX agent

arXiv:2604.23626v1 Announce Type: new Abstract: LLM routing has achieved promising results in integrating the strengths of diverse models while balancing efficiency and performance. However, to suppor

HAC: Parameter-Efficient Hyperbolic Adaptation of CLIP for Zero-Shot VQA

Model ReleasesDGX agent

arXiv:2604.23665v1 Announce Type: new Abstract: Recent advances in representation learning have shown that hyperbolic geometry can offer a more expressive alternative to the Euclidean embeddings used

Hamiltonian Graph Inference Networks: Joint structure discovery and dynamics prediction for lattice Hamiltonian systems from trajectory data

Model ReleasesDGX agent

arXiv:2604.23606v1 Announce Type: new Abstract: Lattice Hamiltonian systems underpin models across condensed matter, nonlinear optics, and biophysics, yet learning their dynamics from data is obstruct

HBGSA: Hydrogen Bond Graph with Self-Attention for Drug-Target Binding Affinity Prediction

Model ReleasesDGX agent

arXiv:2604.23115v1 Announce Type: new Abstract: Accurate prediction of drug-target binding affinity accelerates drug discovery by prioritizing compounds for experimental validation. Current methods fa

Hear from our team:

Model ReleasesDGX agent

Mistral AI shared a message or announcement from their team on X (formerly Twitter), likely featuring team member insights, updates about their AI models or products, or commentary on developments in

Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs

ResearchDGX agent

arXiv:2512.16378v4 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand beyond text, integrating speech as a native modality has given rise to SpeechLLMs, which directly proce

Hidden States Know Where Reasoning Diverges: Credit Assignment via Span-Level Wasserstein Distance

Local AiDGX agent

arXiv:2604.23318v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) performs coarse-grained credit assignment in reinforcement learning with verifiable rewards (RLVR) by assignin

INHerit-SG: Incremental Hierarchical Semantic Scene Graphs with RAG-Style Retrieval

Model ReleasesDGX agent

arXiv:2602.12971v2 Announce Type: replace Abstract: Driven by recent advancements in foundation models, semantic scene graphs have emerged as a promising paradigm for high-level 3D environmental abstr

Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens

SafetyDGX agent

arXiv:2508.01191v5 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has been shown to be effective in eliciting structured reasoning (i.e., CoT reasoning) from large language models (

Learn how to run a local coding agent! Use: - Pi agent - Gemma 4 26B - Serving engine of choice: e.g. LM Studio

Model ReleasesDGX agent

This resource provides instructions for setting up and running a local coding agent using LM Studio's serving engine, featuring the Pi agent framework and Google's Gemma 4 26B language model. It demon

📚 Learn more here https://mistral.ai/news/workflows

Model ReleasesDGX agent

Mistral AI announced new workflow capabilities or features, likely detailing how users can implement multi-step processes or automation using their AI models and services. The announcement was shared

Learning Gradient-based Mixup with Extrapolation toward Flatter Minima for Domain Generalization

Model ReleasesDGX agent

arXiv:2209.14742v2 Announce Type: replace Abstract: To address distribution shifts between training and test data, domain generalization (DG) leverages multiple source domains to learn a model that ge

Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain

Model ReleasesDGX agent

arXiv:2509.10546v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in finance, where unsafe behavior can lead to serious regulatory risks. However, most r

LLMs Reading the Rhythms of Daily Life: Aligned Understanding for Behavior Prediction and Generation

SafetyDGX agent

arXiv:2604.23578v1 Announce Type: cross Abstract: Human daily behavior unfolds as complex sequences shaped by intentions, preferences, and context. Effectively modeling these behaviors is crucial for

Mitigating Error Amplification in Fast Adversarial Training

TutorialsDGX agent

arXiv:2604.24332v1 Announce Type: new Abstract: Fast Adversarial Training (FAT) has proven effective in enhancing model robustness by encouraging networks to learn perturbation-invariant representatio

Multi-View Synergistic Learning with Vision-Language Adaption for Low-Resource Biomedical Image Classification

Model ReleasesDGX agent

arXiv:2604.23977v1 Announce Type: new Abstract: Accurate biomedical image classification under low-resource conditions remains challenging due to limited annotations, subtle inter-class visual differe

MUSIC: Learning Muscle-Driven Dexterous Hand Control

ResearchDGX agent

arXiv:2604.23886v1 Announce Type: cross Abstract: We present a data-driven approach for physics-based, muscle-driven dexterous control that enables musculoskeletal hands to perform precise piano playi

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

Model ReleasesDGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning

Model ReleasesDGX agent

arXiv:2507.06542v4 Announce Type: replace Abstract: Decentralized learning provides a scalable alternative to parameter-server-based training, yet its performance is often hindered by limited peer-to-

Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation

Model ReleasesDGX agent

arXiv:2509.19602v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods have emerged as a promising solution for adapting pre-trained models to various downstream tasks. While thes

Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation

SafetyDGX agent

arXiv:2604.23600v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in persona-driven applications such as education, customer service, and social platforms, where m

Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs

SafetyDGX agent

arXiv:2604.04177v2 Announce Type: replace Abstract: As large language models (LLMs) are increasing integrated into fact-checking pipelines, formal logic is often proposed as a rigorous means by which

Pref-CTRL: Preference Driven LLM Alignment using Representation Editing

Model ReleasesDGX agent

arXiv:2604.23543v1 Announce Type: cross Abstract: Test-time alignment methods offer a promising alternative to fine-tuning by steering the outputs of large language models (LLMs) at inference time wit

Process Supervision of Confidence Margin for Calibrated LLM Reasoning

ResearchDGX agent

arXiv:2604.23333v1 Announce Type: cross Abstract: Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability

Propagation Structure-Semantic Transfer Learning for Robust Fake News Detection

TutorialsDGX agent

arXiv:2604.23974v1 Announce Type: new Abstract: Fake news generally refers to false information that is spread deliberately to deceive people, which has detrimental social effects. Existing fake news

← Previous
1…422423424425426…1059
Next →