AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
19 May 2026

LaunchDarkly launches runtime control layer for the agentic AI era

Model ReleasesDGX agent

LaunchDarkly, a feature control platform that helps developers and software engineers launch and manage products, today announced the launch of AgentControl, a new solution providing real-time managem

LEAF: A Living Benchmark for Event-Augmented Forecasting

Model ReleasesDGX agent

arXiv:2605.16358v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to forecasting. To evaluate this capability while mitigating pre-training data contamination, se

Learned Memory Attenuation in Sage-Husa Kalman Filters for Robust UAV State Estimation

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.18704v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles in dynamic environments face telemetry outages, structural vibrations, and regime-dependent noise that invalidate the station

Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarization

Model ReleasesDGX agent

arXiv:2605.17379v1 Announce Type: cross Abstract: Large language models pretrained on general-domain corpora often exhibit tokenization inefficiencies when applied to specialized domains. Although con

Learning How to Cube

Model ReleasesDGX agent

arXiv:2605.16632v1 Announce Type: cross Abstract: Despite the effectiveness of Cube-and-Conquer (C&C) for solving challenging Boolean Satisfiability (SAT) problems, no prior work has shown that transf

Learning-Zone Energy: Online Data Selection for Efficient RL Post-Training

Model ReleasesDGX agent

arXiv:2605.17003v1 Announce Type: cross Abstract: Reinforcement Learning (RL) post-training has emerged as the dominant paradigm for eliciting mathematical reasoning in Large Language Models (LLMs), y

LERA: LLM-Enhanced RAG for Ad Auction in Generative Chatbots

Model ReleasesDGX agent

arXiv:2605.16474v1 Announce Type: cross Abstract: The integration of advertising auction mechanisms into large language model (LLM)-based chatbots presents a significant opportunity for commercializat

LESSViT: Robust Hyperspectral Representation Learning under Spectral Configuration Shift

Model ReleasesDGX agent

arXiv:2605.18541v1 Announce Type: new Abstract: Modeling hyperspectral imagery (HSI) across different sensors presents a fundamental challenge due to variations in wavelength coverage, band sampling,

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation

Model ReleasesDGX agent

arXiv:2410.13846v3 Announce Type: replace-cross Abstract: Scaling language models to handle longer contexts introduces substantial memory challenges due to the growing cost of key-value (KV) caches. M

Lightweight CNN-Based DDoS Detection for Resource-Constrained Edge Networks

Model ReleasesDGX agent

arXiv:2309.05646v2 Announce Type: replace-cross Abstract: Distributed Denial of Service (DDoS) attacks remain a persistent threat to the availability of Internet services, edge networks, and cyber-phy

LinAlg-Bench: A Forensic Benchmark Revealing Structural Failure Modes in LLM Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2605.16675v1 Announce Type: new Abstract: We introduce LinAlg-Bench, a diagnostic benchmark evaluating 10 frontier large language models on structured linear algebra computation across a strict

LiTS: A Modular Framework for LLM Tree Search

Model ReleasesDGX agent

arXiv:2603.00631v2 Announce Type: replace Abstract: LiTS is a modular Python framework for LLM reasoning via tree search. It decomposes tree search into three reusable components (Policy, Transition,

Live from Code with Claude London: we're launching self-hosted sandboxes (public beta) and MCP tunnels (research preview) in Claude Managed …

Model ReleasesDGX agent

Live from Code with Claude London: we're launching self-hosted sandboxes (public beta) and MCP tunnels (research preview) in Claude Managed Agents. Run agents inside your own perimeter, with your secu

LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injectio

Model ReleasesDGX agent

arXiv:2605.17986v1 Announce Type: cross Abstract: AI agents such as OpenClaw are increasingly deployed in local workflows with access to external tools. This creates indirect prompt-injection (IPI) ri

llm-gemini 0.32

Model ReleasesDGX agent

llm-gemini 0.32 is an alpha release of Simon Willison's LLM Python library and CLI tool that provides access to Google's Gemini models , continuing work on major architectural changes to support newer

llm-gemini 0.32a0

Model ReleasesDGX agent

I don't have current information about this specific entry, so I'll describe what it likely covers based on the available details. This entry documents the release or update of llm-gemini version 0.32

LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models

Model ReleasesDGX agent

arXiv:2605.17653v1 Announce Type: cross Abstract: Sub-billion-parameter Transformer language models are increasingly deployed on edge devices, where the privacy, latency, and operating-cost advantages

LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations

Model ReleasesDGX agent

arXiv:2605.16538v1 Announce Type: cross Abstract: This paper examines the opportunities, limitations, and practical considerations associated with the use of large language models (LLMs) in qualitativ

LongMINT: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems

Model ReleasesDGX agent

arXiv:2605.18565v1 Announce Type: cross Abstract: Real-world agents operate over long and evolving horizons, where information is repeatedly updated and may interfere across memories, requiring accura

LoopQ: Quantization for Recursive Transformers

Model ReleasesDGX agent

arXiv:2605.16343v1 Announce Type: cross Abstract: Looped language models (LoopLMs) improve parameter efficiency by recursively reusing Transformer blocks, enabling deeper computation under a fixed mod

M^2FedAQI: Multimodal Federated Learning for Air Quality Prediction on Heterogeneous Edge Devices

Model ReleasesDGX agent

arXiv:2605.16375v1 Announce Type: new Abstract: Accurate air quality prediction is essential for public health, environmental monitoring, and industrial safety. However, most existing approaches rely

Machine Unlearning for Masked Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.18253v1 Announce Type: cross Abstract: Recent masked diffusion language models (MDLMs), such as LLaDA and Dream, have achieved performance comparable to autoregressive large language models

MADP: A Multi-Agent Pipeline for Sustainable Document Processing with Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2605.17159v1 Announce Type: new Abstract: Document processing automation remains a critical challenge in enterprise environments, where traditional manual approaches are labor-intensive and erro

ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics

Model ReleasesDGX agent

arXiv:2605.18617v1 Announce Type: cross Abstract: Most existing vision-language manipulation research targets rigid robotic arms, whose fixed morphology limits adaptability in cluttered or confined sp

MANTA: Multi-turn Assessment for Nonhuman Thinking & Alignment

Model ReleasesDGX agent

arXiv:2605.16301v1 Announce Type: cross Abstract: Single-turn benchmarks such as AnimalHarmBench (AHB) have established important baselines for measuring animal welfare alignment in large language mod

MARS: Technical Report for the CASTLE Challenge at EgoVis 2026

Model ReleasesDGX agent

arXiv:2605.18176v1 Announce Type: cross Abstract: This report presents MARS, short for Multimodal Agentic Reasoning with Source selection, our system for the CASTLE Challenge at EgoVis 2026. Participa

MAVEN A Multi-Agent Framework for Multicultural Text-to-Video Generation

Model ReleasesDGX agent

arXiv:2605.16716v1 Announce Type: cross Abstract: Text-to-video (T2V) generation has rapidly progressed in visual fidelity, yet its ability to faithfully represent multiple cultures within a single pr

MCQ Difficulty Prediction via Modeling Learner Heterogeneity Using Data-Driven Cognitive Profiling

Model ReleasesDGX agent

arXiv:2605.16290v1 Announce Type: cross Abstract: Predicting the difficulty of multiple-choice questions (MCQs) is important for effective assessment, yet current methods typically assume a unimodal s

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

Model ReleasesDGX agent

arXiv:2603.05308v2 Announce Type: replace-cross Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and claim verification. While large language model

Membership Inference Attacks on Discrete Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.16445v1 Announce Type: cross Abstract: Masked Diffusion Language Models MDLMs replace autoregressive generation with iterative demasking and their privacy properties are largely unstudied.

MemOCR: Layout-Aware Visual Memory for Efficient Long-Horizon Reasoning

Model ReleasesDGX agent

arXiv:2601.21468v5 Announce Type: replace Abstract: Long-horizon agentic reasoning necessitates effectively compressing growing interaction histories into a limited context window. Most existing memor

MentalBench: A DSM-Grounded Benchmark for Evaluating Psychiatric Diagnostic Capability of Large Language Models

Model ReleasesDGX agent

arXiv:2602.12871v2 Announce Type: replace Abstract: Large language models (LLMs) have attracted growing interest as supportive tools for psychiatric assessment and clinical decision support. However,

Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting

Model ReleasesDGX agent

arXiv:2510.10528v3 Announce Type: replace Abstract: Large reasoning models (LRMs) have demonstrated remarkable proficiency in tackling complex tasks through step-by-step thinking. However, this length

MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation

Model ReleasesDGX agent

arXiv:2605.17292v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems have shown promise for solving complex tasks through agent collaboration. However, existing frameworks as

MiniGPT: Rebuilding GPT from First Principles

Model ReleasesDGX agent

arXiv:2605.17398v1 Announce Type: new Abstract: This paper presents MiniGPT, a compact from-scratch implementation of GPT-style autoregressive language modeling in PyTorch. The aim is to rebuild the c

MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery

Model ReleasesDGX agent

arXiv:2605.17198v1 Announce Type: cross Abstract: To be useful for downstream applications, vision decoding models that are trained to reconstruct seen images from human brain activity must be able to

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

Model ReleasesDGX agent

arXiv:2601.08118v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as human simulators, both for evaluating conversational systems and for generating fine-tuning da

Mistral acquires Vienna-based Emmi AI for an undisclosed sum to boost its industrial offerings in Europe; Emmi raised €15M in Austria's largest round in 2025 (Reuters)

Model ReleasesDGX agent

Reuters: Mistral acquires Vienna-based Emmi AI for an undisclosed sum to boost its industrial offerings in Europe; Emmi raised €15M in Austria's largest round in 2025 — Europe's leading artificial int

MixSD: Mixed Contextual Self-Distillation for Knowledge Injection

Model ReleasesDGX agent

arXiv:2605.16865v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to inject new knowledge into language models, but it often degrades pretrained capabilities such as reasonin

Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource

Model ReleasesDGX agent

arXiv:2506.12119v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) language models dramatically expand model capacity and achieve remarkable performance without increasing per-token co

Mixture of Experts for Low-Resource LLMs

Model ReleasesDGX agent

arXiv:2605.17598v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enable efficient model scaling, yet expert routing behavior across underrepresented languages remains poorly unde

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility

Model ReleasesDGX agent

arXiv:2605.16616v1 Announce Type: new Abstract: Autonomous research systems capable of generating complete scientific manuscripts have advanced rapidly, yet robust and realistic evaluation frameworks

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

Model ReleasesDGX agent

arXiv:2602.22667v2 Announce Type: replace Abstract: Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abunda

MorphSeek: Fine-grained Latent Representation-Level Policy Optimization for Deformable Image Registration

Model ReleasesDGX agent

arXiv:2511.17392v3 Announce Type: replace Abstract: Deformable image registration (DIR) remains a fundamental yet challenging problem in medical image analysis, largely due to the prohibitively high-d

Multi-Party Multi-Objective Optimization as Consensus Search: Runtime Analysis of Cross-Party Recombination

Model ReleasesDGX agent

arXiv:2605.17454v1 Announce Type: new Abstract: Multi-party multi-objective optimization problems (MPMOPs) require consensus among autonomous decision makers and therefore differ from flattened many-o

Multi-site PPG: An In-the-Wild Physiological Dataset from Emerging Multi-site Wearables

Model ReleasesDGX agent

arXiv:2605.17859v1 Announce Type: cross Abstract: Wearables are widely used for mobile health monitoring, and photoplethysmography (PPG) is a key sensing modality for heart rate and related physiologi

Multilingual jailbreaking of LLMs using low-resource languages

Model ReleasesDGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.16409v1 Announce Type: cross Abstract: Optical character recognition (OCR) and multilingual text understanding remain major failure modes of multimodal large language models (MLLMs), partic

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

Model ReleasesDGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

My notes on Gemini 3.5 Flash - 3x the price of Gemini 3 Flash but Google are planning to use it for many of their own products https://simon…

Model ReleasesDGX agent

Google's Gemini 3.5 Flash model costs approximately 3x more than Gemini 3 Flash, despite being a newer version. Google plans to integrate Gemini 3.5 Flash into many of their own products, suggesting t

NeuroMAS: Multi-Agent Systems as Neural Networks with Joint Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.16757v1 Announce Type: new Abstract: Multi-agent language systems are often built as hand-designed workflows, where agents are assigned semantic roles and communication protocols are specif

Neuroscience-inspired Staged Representation Learning with Disentangled Coarse- and Fine-Grained Semantics for EEG Visual Decoding

Model ReleasesDGX agent

arXiv:2605.16923v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) signals remains a fundamental challenge in brain-computer interfaces and medical rehabilit

NeuSymMS: A Hybrid Neuro-Symbolic Memory System for Persistent, Self-Curating LLM Agents

Model ReleasesDGX agent

arXiv:2605.17596v1 Announce Type: new Abstract: We present NeuSymMS, an adaptive memory system that enables large language model (LLM) agents to learn, remember, and reason about users across sessions

New upgrades to the @GeminiApp are you helping you get more done: ✨Gemini Spark is your 24/7 personal AI agent that can take action on your …

Model ReleasesDGX agent

New upgrades to the @GeminiApp are you helping you get more done: ✨Gemini Spark is your 24/7 personal AI agent that can take action on your behalf, under your direction. It seamlessly integrates with

NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation

Model ReleasesDGX agent

arXiv:2605.17364v1 Announce Type: new Abstract: Media bias detection has predominantly been framed as a classification task: assign a political label to an article or outlet. We argue this framing is

Noise2Params: Unification and Parameter Determination from Noise via a Probabilistic Event Camera Model

Model ReleasesDGX agent

arXiv:2605.16317v1 Announce Type: new Abstract: Accurate, unified models for event cameras (ECs) remain elusive, hampering calibration and algorithm design. We develop a foundational probabilistic mod

Nori Bot: A Sub-$1,000 Floor-to-Counter Mobile Manipulator

Model ReleasesDGX agent

arXiv:2605.16537v1 Announce Type: new Abstract: Open-source mobile manipulators have reached 660 (XLeRobot) but every sub-1,000 platform shares three limitations: a fixed-height workspace, reactive-on

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

Model ReleasesDGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

Now also on the Claude Blog: https://claude.com/blog/using-claude-code-the-unreasonable-effectiveness-of-html

Model ReleasesDGX agent

This blog post from Anthropic's Claude discusses the practical effectiveness of using HTML within Claude's code interpreter, likely exploring how HTML can be leveraged for web development, data visual

Offline Contextual Bandits in the Presence of New Actions

Model ReleasesDGX agent

arXiv:2605.18509v1 Announce Type: new Abstract: Automated decision-making algorithms drive applications such as recommendation systems and search engines. These algorithms often rely on off-policy con

← Previous
1…236237238239240…377
Next →