AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,952 results
24 Apr 2026

Weighting What Matters: Boosting Sample Efficiency in Medical Report Generation via Token Reweighting

ResearchDGX agent

arXiv:2604.21082v1 Announce Type: new Abstract: Training vision-language models (VLMs) for medical report generation is often hindered by the scarcity of high-quality annotated data. This work evaluat

23 Apr 2026

A pelican for GPT-5.5 via the semi-official Codex backdoor API

Model ReleasesDGX agent

GPT-5.5 is out. It's available in OpenAI Codex and is rolling out to paid ChatGPT subscribers. I've had some preview access and found it to be a fast, effective and highly capable model. As is usually

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2604.20273v1 Announce Type: new Abstract: We present ActuBench, a multi-agent LLM pipeline for the automated generation and evaluation of advanced actuarial assessment items aligned with the Int

Always enjoy getting to chat with @swyx on our annual cross-episode with @latentspacepod on the state of AI. We hit on what’s shifted, what …

Model ReleasesDGX agent

Always enjoy getting to chat with @swyx on our annual cross-episode with @latentspacepod on the state of AI. We hit on what’s shifted, what surprised us and what’s next. We covered: ▪️ Whether AI infr

Anchor-and-Resume Concession Under Dynamic Pricing for LLM-Augmented Freight Negotiation

Model ReleasesDGX agent

arXiv:2604.20732v1 Announce Type: cross Abstract: Freight brokerages negotiate thousands of carrier rates daily under dynamic pricing conditions where models frequently revise targets mid-conversation

Anthropic’s Mythos breach was humiliating

Model ReleasesDGX agent

Anthropic's tightly controlled rollout of Claude Mythos has taken an awkward turn. After spending weeks insisting the AI model is so capable at cybersecurity that it is too dangerous to release public

ATIR: Towards Audio-Text Interleaved Contextual Retrieval

Model ReleasesDGX agent

arXiv:2604.20267v1 Announce Type: cross Abstract: Audio carries richer information than text, including emotion, speaker traits, and environmental context, while also enabling lower-latency processing

Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Control

Model ReleasesDGX agent

arXiv:2601.02896v2 Announce Type: replace Abstract: Controlling emergent behavioral personas (e.g., sycophancy, hallucination) in Large Language Models (LLMs) is critical for AI safety, yet remains a

Camera Control for Text-to-Image Generation via Learning Viewpoint Tokens

TutorialsDGX agent

arXiv:2604.19954v1 Announce Type: new Abstract: Current text-to-image models struggle to provide precise camera control using natural language alone. In this work, we present a framework for precise c

Catalyzing Informed Residential Energy Retrofit Decisions via Domain-Specific LLM

Model ReleasesDGX agent

arXiv:2602.20181v2 Announce Type: replace-cross Abstract: Residential energy retrofit initiation is often stalled by an expertise gap, where homeowners lack the technical literacy required for structu

Claude is connecting directly to your personal apps like Spotify, Uber Eats, and TurboTax

Model ReleasesDGX agent

Claude users can access more apps with Anthropic's AI now thanks to new connectors for everything from hiking to grocery shopping. Anthropic already supported connecting numerous work-related apps to

CoAuthorAI: A Human in the Loop System For Scientific Book Writing

ResearchDGX agent

arXiv:2604.19772v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in scientific writing but struggle with book-length tasks, often producing inconsistent structure a

CXR-LanIC: Language-Grounded Interpretable Classifier for Chest X-Ray Diagnosis

ResearchDGX agent

arXiv:2510.21464v2 Announce Type: replace Abstract: Deep learning models have achieved remarkable accuracy in chest X-ray diagnosis, yet their widespread clinical adoption remains limited by the black

Databricks partners with OpenAI on GPT-5.5

Model ReleasesDGX agent

Databricks has announced a partnership with OpenAI to integrate GPT-5.5, OpenAI's latest language model, into its data and AI platform. This collaboration enables Databricks customers to leverage GPT-

Day 0 vLLM support for Qwen3.6-27B! @vllm_project ♥️❤️

Model ReleasesDGX agent

Day 0 vLLM support for Qwen3.6-27B! @vllm_project ♥️❤️ 🎉 Day-0 vLLM support for Qwen3.6-27B! Congrats to @Alibaba_Qwen on the new 27B dense model release. Looking forward to more of the Qwen3.6 series

Day 2 at Google Cloud Next: A marathon developer keynote

Model ReleasesDGX agent

At Google Cloud, every day is Developer Day, but none so much as day 2 of Google Cloud Next, when we hold the developer keynote. This year’s topic? An in-depth look at Gemini Enterprise Agent Platform

DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation

AgentsDGX agent

arXiv:2604.20841v1 Announce Type: new Abstract: Recent advances in video generative models enable the synthesis of realistic human-object interaction videos across a wide range of scenarios and object

Dual-Cluster Memory Agent: Resolving Multi-Paradigm Ambiguity in Optimization Problem Solving

AgentsDGX agent

arXiv:2604.20183v1 Announce Type: new Abstract: Large Language Models (LLMs) often struggle with structural ambiguity in optimization problems, where a single problem admits multiple related but confl

Efficient Reinforcement Learning using Linear Koopman Dynamics for Nonlinear Robotic Systems

SafetyDGX agent

arXiv:2604.19980v1 Announce Type: new Abstract: This paper presents a model-based reinforcement learning (RL) framework for optimal closed-loop control of nonlinear robotic systems. The proposed appro

Excretion Detection in Pigsties Using Convolutional and Transformerbased Deep Neural Networks

ResearchDGX agent

arXiv:2412.00256v3 Announce Type: replace Abstract: Animal excretions in form of urine puddles and feces are a significant source of emissions in livestock farming. Automated detection of soiled floor

FlashNorm: Fast Normalization for Transformers

Model ReleasesDGX agent

arXiv:2407.09577v4 Announce Type: replace Abstract: Normalization layers are ubiquitous in large language models (LLMs) yet represent a compute bottleneck: on hardware with distinct vector and matrix

From Competition to Synergy: Unlocking Reinforcement Learning for Subject-Driven Image Generation

ResearchDGX agent

arXiv:2510.18263v2 Announce Type: replace-cross Abstract: Subject-driven image generation models face a fundamental trade-off between identity preservation (fidelity) and prompt adherence (editability

How to Use Transformers.js in a Chrome Extension

TutorialsDGX agent

This guide explains how to integrate Transformers.js, a JavaScript library for running machine learning models, into Chrome extensions to enable on-device AI capabilities. It covers the technical setu

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror

Model ReleasesDGX agent

arXiv:2604.14785v2 Announce Type: replace Abstract: Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their poten

'Newspaper Eat' Means 'Not Tasty': A Taxonomy and Benchmark for Coded Language in Real-World Chinese Online Reviews

Model ReleasesDGX agent

arXiv:2601.19932v2 Announce Type: replace Abstract: Coded language is an important part of human communication. It refers to cases where users intentionally encode meaning so that the surface text dif

OpenAI releases GPT-5.5 with advanced math, coding capabilities

Model ReleasesDGX agent

OpenAI Group PBC today launched a new large language model that is significantly better than its predecessors at solving math problems and writing code. GPT-5.5 is rolling out a week after rival Anthr

Option Pricing on Noisy Intermediate-Scale Quantum Computers: A Quantum Neural Network Approach

Model ReleasesDGX agent

arXiv:2604.19832v1 Announce Type: cross Abstract: In a global derivatives market with notional values in the hundreds of trillions of dollars, the accuracy and efficiency of pricing models are of fund

ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability

Model ReleasesDGX agent

arXiv:2508.07050v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) based listwise ranking has shown superior performance in many passage ranking tasks. With the development of Large

SurgCoT: Advancing Spatiotemporal Reasoning in Surgical Videos through a Chain-of-Thought Benchmark

Model ReleasesDGX agent

arXiv:2604.20319v1 Announce Type: new Abstract: Fine-grained spatiotemporal reasoning on surgical videos is critical, yet the capabilities of Multi-modal Large Language Models (MLLMs) in this domain r

Survival of the Cheapest: Cost-Aware Hardware Adaptation for Adversarial Robustness

Model ReleasesDGX agent

arXiv:2409.07609v2 Announce Type: replace-cross Abstract: Deploying adversarially robust machine learning systems requires continuous trade-offs between robustness, cost, and latency. We present an au

The Expense of Seeing: Attaining Trustworthy Multimodal Reasoning Within the Monolithic Paradigm

ResearchDGX agent

arXiv:2604.20665v1 Announce Type: cross Abstract: The rapid proliferation of Vision-Language Models (VLMs) is widely celebrated as the dawn of unified multimodal knowledge discovery but its foundation

The Optical and Infrared Are Connected

ResearchDGX agent

arXiv:2503.03816v2 Announce Type: replace-cross Abstract: Galaxies are often modelled as composites of separable components with distinct spectral signatures, implying that different wavelength ranges

Tokenised Flow Matching for Hierarchical Simulation Based Inference

Model ReleasesDGX agent

arXiv:2604.20723v1 Announce Type: cross Abstract: The cost of simulator evaluations is a key practical bottleneck for Simulation Based Inference (SBI). In hierarchical settings with shared global para

Transparent Screening for LLM Inference and Training Impacts

ResearchDGX agent

arXiv:2604.19757v1 Announce Type: cross Abstract: This paper presents a transparent screening framework for estimating inference and training impacts of current large language models under limited obs

UCCL-Zip: Lossless Compression Supercharged GPU Communication

Model ReleasesDGX agent

arXiv:2604.17172v2 Announce Type: replace-cross Abstract: The rapid growth of large language models (LLMs) has made GPU communication a critical bottleneck. While prior work reduces communication volu

veScale-FSDP: Flexible and High-Performance FSDP at Scale

ResearchDGX agent

arXiv:2602.22437v3 Announce Type: replace-cross Abstract: Fully Sharded Data Parallel (FSDP), also known as Zero Redundancy Optimizer (ZeRO), is widely used for large-scale model training, because of

Video-ToC: Video Tree-of-Cue Reasoning

Model ReleasesDGX agent

arXiv:2604.20473v1 Announce Type: new Abstract: Existing Video Large Language Models (Video LLMs) struggle with complex video understanding, exhibiting limited reasoning capabilities and potential hal

Where are they looking in the operating room?

ResearchDGX agent

arXiv:2604.20574v1 Announce Type: new Abstract: Purpose: Gaze-following, the task of inferring where individuals are looking, has been widely studied in computer vision, advancing research in visual a

22 Apr 2026

[AINews] OpenAI launches GPT-Image-2

Model ReleasesDGX agent

OpenAI has launched GPT-Image-2, an advancement in their image generation capabilities. The model likely represents improvements over previous versions in areas such as image quality, prompt understan

Analytical Extraction of Conditional Sobol' Indices via Basis Decomposition of Polynomial Chaos Expansions

Model ReleasesDGX agent

arXiv:2604.19165v1 Announce Type: cross Abstract: In uncertainty quantification, evaluating sensitivity measures under specific conditions (i.e., conditional Sobol' indices) is essential for systems w

Assessing VLM-Driven Semantic-Affordance Inference for Non-Humanoid Robot Morphologies

SafetyDGX agent

arXiv:2604.19509v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in understanding human-object interactions, but their application to robotic sys

BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps

ResearchDGX agent

arXiv:2604.19532v1 Announce Type: cross Abstract: Tokenizing music to fit the general framework of language models is a compelling challenge, especially considering the diverse symbolic structures in

CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation

Model ReleasesDGX agent

arXiv:2602.20409v2 Announce Type: replace Abstract: Recent vision-language models (VLMs) such as CLIP demonstrate impressive cross-modal reasoning, extending beyond images to 3D perception. Yet, these

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks

Model ReleasesDGX agent

arXiv:2604.19262v1 Announce Type: cross Abstract: Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities.

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?

Model ReleasesDGX agent

arXiv:2602.18571v2 Announce Type: replace-cross Abstract: While significant progress has been made in automating various aspects of software development through coding agents, there is still significa

Decoupled DiLoCo: A new frontier for resilient, distributed AI training

Model ReleasesDGX agent

Decoupled DiLoCo is a distributed architecture that enables training of large language models across distant data centers using lower bandwidth and improved hardware resilience by dividing training in

Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image Generation

SafetyDGX agent

arXiv:2604.19141v1 Announce Type: new Abstract: Diffusion- and flow-based models usually allocate compute uniformly across space, updating all patches with the same timestep and number of function eva

Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps

Model ReleasesDGX agent

arXiv:2604.19565v1 Announce Type: cross Abstract: Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on gold-standard out

Detoxification for LLM: From Dataset Itself

Local AiDGX agent

arXiv:2604.19124v1 Announce Type: new Abstract: Existing detoxification methods for large language models mainly focus on post-training stage or inference time, while few tackle the source of toxicity

Discrete Tilt Matching

ResearchDGX agent

arXiv:2604.18739v1 Announce Type: new Abstract: Masked diffusion large language models (dLLMs) are a promising alternative to autoregressive generation. While reinforcement learning (RL) methods have

Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture The Flag Challenges

Model ReleasesDGX agent

arXiv:2604.19354v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly proposed for autonomous cybersecurity tasks, but their capabilities in realistic offensive settings r

Energy-Weighted Flow Matching: Unlocking Continuous Normalizing Flows for Efficient and Scalable Boltzmann Sampling

Model ReleasesDGX agent

arXiv:2509.03726v2 Announce Type: replace-cross Abstract: Sampling from unnormalized target distributions, e.g. Boltzmann distributions mu_{ext{target}}(x) propto exp(-E(x)/T), is fundamental to many

FB-NLL: A Feature-Based Approach to Tackle Noisy Labels in Personalized Federated Learning

SafetyDGX agent

arXiv:2604.19729v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) aims to learn multiple task-specific models rather than a single global model across heterogeneous data distributi

FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion

Model ReleasesDGX agent

arXiv:2604.19015v1 Announce Type: cross Abstract: Federated fine-tuning of Large Language Models (LLMs) is obstructed by a trilemma of challenges: protecting LLMs intellectual property (IP), ensuring

GenerativeMPC: VLM-RAG-guided Whole-Body MPC with Virtual Impedance for Bimanual Mobile Manipulation

Model ReleasesDGX agent

arXiv:2604.19522v1 Announce Type: new Abstract: Bimanual mobile manipulation requires a seamless integration between high-level semantic reasoning and safe, compliant physical interaction - a challeng

Gradient-Based Program Synthesis with Neurally Interpreted Languages

TutorialsDGX agent

arXiv:2604.18907v1 Announce Type: cross Abstract: A central challenge in program induction has long been the trade-off between symbolic and neural approaches. Symbolic methods offer compositional gene

HELM: Harness-Enhanced Long-horizon Memory for Vision-Language-Action Manipulation

Model ReleasesDGX agent

arXiv:2604.18791v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models fail systematically on long-horizon manipulation tasks despite strong short-horizon performance. We show that this

HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing

Model ReleasesDGX agent

arXiv:2604.19071v1 Announce Type: new Abstract: Evaluating the writing capabilities of large language models (LLMs) remains a significant challenge due to the multidimensional nature of writing skills

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships …

Model ReleasesDGX agent

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships SOTA models ] This isn't just another ML Research Loop wrapp

Improvements to the post-processing of weather forecasts using machine learning and feature selection

ResearchDGX agent

arXiv:2604.19340v1 Announce Type: cross Abstract: This study aims to develop and improve machine learning-based post-processing models for precipitation, temperature, and wind speed predictions using

← Previous
1…364365366367368…1050
Next →