AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
2 Jun 2026

Cross-lingual Self-Consistency for Multilingual Reasoning with Language Models

ResearchDGX agent

arXiv:2606.01464v1 Announce Type: new Abstract: Despite expanding their multilingual coverage, the advanced reasoning capabilities of LLMs remain largely confined to a few high-resource languages like

CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards

Model ReleasesDGX agent

arXiv:2606.00020v1 Announce Type: cross Abstract: Large Language Model (LLM) based Chinese Grammatical Error Correction (CGEC) systems face two critical challenges: general-purpose models lack special

DAG-MoE: From Simple Mixture to Structural Aggregation in Mixture-of-Experts

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.01062v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become a leading approach for decoupling parameter count from computational cost in large language models, yet effe

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark

Model ReleasesDGX agent

arXiv:2606.02214v1 Announce Type: new Abstract: Large language models are increasingly used in value-sensitive decision settings, where irrelevant demographic cues should not alter judgments. We const

Drifting Preference Optimization for One-Step Generative Models

SafetyDGX agent

arXiv:2606.02521v1 Announce Type: cross Abstract: One-step text-to-image generators are attractive for deployment because they generate an image with a single forward pass, but preference finetuning t

Efficient Exploration for Iterative Nash Preference Optimization

Model ReleasesDGX agent

arXiv:2606.01382v1 Announce Type: cross Abstract: Preference alignment is central to improving large language models, but standard reward-based formulations can be restrictive when human preferences a

Efficient Weighted Sampling via Score-based Generative Models

ResearchDGX agent

arXiv:2502.04646v2 Announce Type: replace-cross Abstract: Weighted sampling -- sampling from a probability density function (PDF) proportional to the product of a base PDF and a weight function -- is

Enhancing LLM Metacognition via Cognitive Pairwise Training

Model ReleasesDGX agent

arXiv:2606.00869v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to LLM reasoning, but its outcome-level rewards can make models more willing to

Evolving to the Aesthetics of a Vision-Language Model

ApplicationsDGX agent

arXiv:2606.00112v1 Announce Type: cross Abstract: Evolutionary systems have demonstrated remarkable results in creative domains, with recent applications in generative typography, design, and music. H

Experimenting with TPUs, GKE Managed DRANET, and Multi-cluster Inference Gateway

Model ReleasesDGX agent

What happens when your workload fails in one region but you need access to service? This is a common case for availability and uptime. With recent enhancement to the Kubernetes ecosystem and capabilit

FedMTFI: Feature Importance Based Optimized Multi Teacher Knowledge Distillation in Heterogeneous Federated Learning Environment

Local AiDGX agent

arXiv:2606.01607v1 Announce Type: cross Abstract: Federated learning (FL) is a decentralized approach that enables collaborative model training without exposing raw data. Instead of transferring sensi

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning

Model ReleasesDGX agent

arXiv:2604.03893v2 Announce Type: replace Abstract: Current multimodal benchmarks for scientific reasoning primarily evaluate local information extraction -- models recognize symbols and values and th

FlatVPR: Plug-and-play Geo-linear Residual Adapter for Geometric Rectification of Foundation Model Feature Manifolds

ResearchDGX agent

arXiv:2606.01734v1 Announce Type: new Abstract: This paper proposes ``FlatVPR,'' a novel geometric rectification paradigm that effectively bridges the trade-off between map lightweightness and localiz

HypothesisMed: Inference-Time Answer Fusion and Structured Hypothesis-Space Reporting for Biomedical Question Answering

Model ReleasesDGX agent

arXiv:2606.00971v1 Announce Type: new Abstract: Biomedical question answering with large language models is commonly evaluated using answer accuracy, but answer accuracy alone does not indicate whethe

LAP: Fast LAtent Diffusion Planner for Autonomous Driving

Model ReleasesDGX agent

arXiv:2512.00470v4 Announce Type: replace Abstract: Diffusion models have demonstrated strong capabilities for modeling human-like driving behaviors in autonomous driving, but their iterative sampling

Large Language Model Guided Incentive Aware Reward Design for Cooperative Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2603.24324v4 Announce Type: replace-cross Abstract: Designing effective auxiliary rewards for cooperative multi-agent systems remains challenging, as misaligned incentives can induce suboptimal

Microsoft’s first advanced reasoning AI is here

IndustryDGX agent

Microsoft announced a bunch of new in-house AI models at Build 2026, including a new 'flagship' model: MAI-Thinking-1. It's an ambitious step into model development for Microsoft, which introduced its

Modeling Spectral Energy Shifts in Spatio-Temporal Graph Anomaly Detection

ResearchDGX agent

arXiv:2606.00304v1 Announce Type: new Abstract: Graph anomaly detection methods aim to distinguish anomalous nodes. While prior methods characterize anomalies through increased variation in the spectr

Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events

Model ReleasesDGX agent

arXiv:2606.02522v1 Announce Type: cross Abstract: Video multimodal large language models (MLLMs) have made rapid progress on general and long-form video understanding, yet their ability to preserve br

MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition

SafetyDGX agent

arXiv:2606.00708v1 Announce Type: new Abstract: Automated data science is a structured model-selection problem. A solution must choose data transformations, feature representations, architecture, trai

MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching

Model ReleasesDGX agent

arXiv:2606.01985v1 Announce Type: new Abstract: Recent breakthroughs in instruction-based image editing have captured significant attention, as models are now capable of handling real-world editing de

Nonlinear Equilibrium Transitions in a Potential Game Model for Federated Learning

ResearchDGX agent

arXiv:2411.11793v2 Announce Type: replace Abstract: In federated learning (FL), a central server typically allocates training efforts to clients. However, from a market-oriented perspective, clients m

OncoReason: Structuring Clinical Reasoning in LLMs for Robust and Interpretable Survival Prediction

Model ReleasesDGX agent

arXiv:2510.17532v2 Announce Type: replace Abstract: Predicting cancer treatment outcomes requires models that are both accurate and interpretable, particularly in the presence of heterogeneous clinica

Paving the Way for Point Cloud Video Representation Learning Using A PDE Model

ResearchDGX agent

arXiv:2606.01604v1 Announce Type: new Abstract: Investigating spatial-temporal correlations, specifically how spatial points vary over time, is crucial for understanding point cloud videos. Traditiona

Preconditioned One-Step Generative Modeling for Bayesian Inverse Problems in Function Spaces

ResearchDGX agent

arXiv:2603.14798v2 Announce Type: replace-cross Abstract: We propose a machine-learning algorithm for Bayesian inverse problems in the function-space regime. Based on one-step generative transport, th

Residual Decoder Adapter: ID-Preserving Tokenizer Adaption for Autoregressive Text Rendering

Model ReleasesDGX agent

arXiv:2606.01911v1 Announce Type: new Abstract: Visual Autoregressive (AR) models generate images by predicting discrete tokens that are decoded by a visual tokenizer. Despite demonstrating strong ove

Safe2Drive: Evaluating Safe Driving Behaviors of E2E Autonomous Driving Models

SafetyDGX agent

arXiv:2606.00191v1 Announce Type: cross Abstract: Recent end-to-end (E2E) autonomous driving policies achieve high driving scores in closed-loop simulations. Yet it remains unclear whether these polic

SindBERT, the Sailor: Charting the Seas of Turkish NLP

Model ReleasesDGX agent

arXiv:2510.21364v2 Announce Type: replace Abstract: Transformer models have revolutionized NLP, yet many morphologically rich languages remain underrepresented in large-scale pre-training efforts. Wit

Snowflake adds new AI services while continuing to build relationships with key model providers

IndustryDGX agent

In the era of artificial intelligence, some companies have struggled to adopt artificial intelligence and others have pivoted to an AI framework that has yielded positive results. Snowflake Inc. this

Spatial Representation Learning Beyond Pixels: Unifying Raster Data and Vector Semantics for Human-Centric Geospatial Foundation Models

TutorialsDGX agent

arXiv:2606.02374v1 Announce Type: new Abstract: Earth Observation (EO) has fundamentally transformed the monitoring of environmental processes and human activities up to planetary scale. Recent advanc

Structure and Scale in Simplicial Sequence Modelling

ResearchDGX agent

arXiv:2606.01302v1 Announce Type: new Abstract: Modern large-scale deep learning exhibits two striking empirical phenomena: behavioural scaling laws (predictable performance gains with increasing scal

The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete

Model ReleasesDGX agent

arXiv:2606.00048v1 Announce Type: cross Abstract: Prior research has established that instruction-tuned large language models exhibit left-of-center political bias, measured exclusively through abstra

TimeBlocks: Foundational and Continual Time-Series Blockbase -- Extended Version

ResearchDGX agent

arXiv:2606.02142v1 Announce Type: new Abstract: The ongoing digitization has led to a proliferation of time-series data streams that monitor a variety of processes, from which valuable insights may be

Truth, Trust, and Trouble: Medical AI on the Edge

Model ReleasesDGX agent

arXiv:2507.02983v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) hold significant promise for transforming digital health by enabling automated medical question answering. Howeve

Video Reasoning without Training

TutorialsDGX agent

arXiv:2510.17045v2 Announce Type: replace-cross Abstract: Video reasoning using Large Multimodal Models (LMMs) relies on costly reinforcement learning (RL) and verbose chain-of-thought, resulting in s

WAXAL-NET: Finetuned Edge ASR Across 19 African Languages

ResearchDGX agent

arXiv:2606.02375v1 Announce Type: new Abstract: We evaluate whether compact domain-specialized ASR models can outperform massively multilingual foundation models for conversational African speech acro

Zamba2-VL Technical Report

Model ReleasesDGX agent

arXiv:2606.00390v1 Announce Type: cross Abstract: We present Zamba2-VL, a suite of vision-language models built on Zamba2, a hybrid language-model architecture combining Mamba2 state-space layers with

1 Jun 2026

A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents

AgentsDGX agent

arXiv:2602.08964v2 Announce Type: replace-cross Abstract: Understanding an agent's goals helps explain and predict its behaviour, yet there is no established methodology for reliably attributing goals

A Lecture Note on Offline RL and IRL, Part II: Foundations of Inverse Reinforcement Learning and Dynamic Discrete Choice Models

SafetyDGX agent

arXiv:2605.30843v1 Announce Type: new Abstract: In the forward reinforcement-learning problem, the reward is fixed and known; the learner is asked to find a good policy or value function. Here we turn

A Pilot Study on Curator-Guided Multilingual Art Description for Blind and Low-Vision Audiences with Small Vision-Language Models

ResearchDGX agent

arXiv:2605.31080v1 Announce Type: cross Abstract: Blind and low-vision (BLV) audiences remain underserved by visual art descriptions, particularly across languages and in museum settings where privacy

CobSeg: Coherence Boundary Modeling for Dialogue Topic Segmentation

Local AiDGX agent

arXiv:2605.30668v1 Announce Type: cross Abstract: Dialogue topic segmentation is critical in many human-AI collaborative applications which requires identifying heterogeneous boundary cues, including

Count Anything

Model ReleasesDGX agent

arXiv:2605.30846v1 Announce Type: new Abstract: Object counting remains fragmented across domain-specific datasets and task formulations, despite rapid progress in generalist vision models. Existing c

GPU Forecasters: Language Models as Selective Surrogates for Kernel Runtime Optimization

Local AiDGX agent

arXiv:2605.31464v1 Announce Type: cross Abstract: GPU kernels are the workhorse of modern deep learning, and optimizing them (via evolutionary search or coding agents) usually requires repeated measur

Hedging on the Frontier: Learning New Tasks with Few Samples

Model ReleasesDGX agent

arXiv:2605.30997v1 Announce Type: cross Abstract: When a learner faces a new task with few samples, it must leverage any available side information. In practice, this often comes in the form of model

HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

SafetyDGX agent

arXiv:2602.16165v2 Announce Type: replace-cross Abstract: Training LLMs as interactive agents for multi-turn decision-making remains challenging, particularly in long-horizon tasks with sparse and del

Import AI 459: AI oversight is difficult; scaling laws for protein folding models; and pricing the extinction risk of AI systems

SafetyDGX agent

This newsletter issue discusses three key topics in AI development and safety: the challenges involved in overseeing and controlling advanced AI systems, empirical findings about how protein folding A

Knowledge Boundary Probing and Demand-Guided Intervention for LLM-Based Power System Code Generation

Model ReleasesDGX agent

arXiv:2605.31478v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to automate power-system analysis, but many utilities and energy-research labs require on-premise s

Object-Informed Model Predictive Path Integral Control for Non-Prehensile Robot Manipulation

ResearchDGX agent

arXiv:2605.30778v1 Announce Type: new Abstract: Long-horizon planning for non-prehensile robot manipulation is challenging due to underactuated and discontinuous interactions. We propose a hierarchica

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

Model ReleasesDGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

Query-focused and Memory-aware Reranker for Long Context Processing

Model ReleasesDGX agent

arXiv:2602.12192v3 Announce Type: replace Abstract: Built upon the existing analysis of retrieval heads in large language models, we propose an alternative reranking framework that trains models to es

Rethinking Efficient Crack Segmentation with Task-Aligned Structural-Directional Modeling

Local AiDGX agent

arXiv:2605.31048v1 Announce Type: new Abstract: Recent crack segmentation methods often follow generic semantic segmentation designs, using stronger backbones, hybrid CNN-Transformer-Mamba encoders, a

Scaling Multi-Agent Environment Co-Design with Diffusion Models

SafetyDGX agent

arXiv:2511.03100v2 Announce Type: replace-cross Abstract: The agent-environment co-design paradigm jointly optimises agent policies and environment configurations in search of improved system performa

Steering LLMs? Actually, Sparse Autoencoders can outperform simple baselines

Model ReleasesDGX agent

arXiv:2605.31183v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) have been seen as a promising avenue for exploring the internals of Large Language Models (LLMs) and for steering model out

VLM-GLoc: Vision-Language Model Enhanced Monte Carlo Localization for Robust Semantic Global Localization in Cluttered Quasi-Static Environments

ApplicationsDGX agent

arXiv:2605.30506v1 Announce Type: cross Abstract: Global localization in geometrically aliased, quasi-static environments such as grocery stores, offices, schools, and hospitals poses a significant ch

31 May 2026

What’s new in Microsoft Foundry | May 2026

Model ReleasesDGX agent

May ships trace-based evaluation for any agent on any cloud, Grok 4.3 and DeepSeek V4 in the model catalog, GPT-5 Reinforcement Fine-Tuning at gated GA, three Microsoft Research on-device agent models

29 May 2026

A Dual-Path Architecture for Scaling Compute and Capacity in LLMs

Model ReleasesDGX agent

arXiv:2605.30202v1 Announce Type: new Abstract: Looped transformers apply a shared block multiple times and have emerged as a parameter-efficient route to scaling compute in language models. However,

A Matter of Interest: Understanding Interestingness of Math Problems in Humans and Language Models

ApplicationsDGX agent

arXiv:2511.08548v2 Announce Type: replace Abstract: The evolution of mathematics is shaped importantly by interestingness: researchers choose which problems to pursue, and students choose which proble

AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling

ResearchDGX agent

arXiv:2605.29488v1 Announce Type: cross Abstract: Conditional human motion generation remains a fundamental challenge in computer vision and robotics. Despite significant progress, current methods are

AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligence

Model ReleasesDGX agent

arXiv:2605.21739v2 Announce Type: replace Abstract: Emotional intelligence (EI), the ability to perceive, understand, and respond appropriately to others' emotional states, is central to human communi

BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices

AgentsDGX agent

arXiv:2605.29705v1 Announce Type: new Abstract: Trajectory prediction is a fundamental task for autonomous systems, requiring complex reasoning about multi-agent interactions and intents. Large langua

← Previous
1…257258259260261…1034
Next →