AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,611 results
Model Releases

Learned Memory Attenuation in Sage-Husa Kalman Filters for Robust UAV State Estimation

DGX agent

arXiv:2605.18704v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles in dynamic environments face telemetry outages, structural vibrations, and regime-dependent noise that invalidate the station

model-releasesarxiv-cs-lg
19 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarization

DGX agent

arXiv:2605.17379v1 Announce Type: cross Abstract: Large language models pretrained on general-domain corpora often exhibit tokenization inefficiencies when applied to specialized domains. Although con

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Learning How to Cube

DGX agent

arXiv:2605.16632v1 Announce Type: cross Abstract: Despite the effectiveness of Cube-and-Conquer (C&C) for solving challenging Boolean Satisfiability (SAT) problems, no prior work has shown that transf

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Learning-Zone Energy: Online Data Selection for Efficient RL Post-Training

DGX agent

arXiv:2605.17003v1 Announce Type: cross Abstract: Reinforcement Learning (RL) post-training has emerged as the dominant paradigm for eliciting mathematical reasoning in Large Language Models (LLMs), y

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LERA: LLM-Enhanced RAG for Ad Auction in Generative Chatbots

DGX agent

arXiv:2605.16474v1 Announce Type: cross Abstract: The integration of advertising auction mechanisms into large language model (LLM)-based chatbots presents a significant opportunity for commercializat

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LESSViT: Robust Hyperspectral Representation Learning under Spectral Configuration Shift

DGX agent

arXiv:2605.18541v1 Announce Type: new Abstract: Modeling hyperspectral imagery (HSI) across different sensors presents a fundamental challenge due to variations in wavelength coverage, band sampling,

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation

DGX agent

arXiv:2410.13846v3 Announce Type: replace-cross Abstract: Scaling language models to handle longer contexts introduces substantial memory challenges due to the growing cost of key-value (KV) caches. M

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Lightweight CNN-Based DDoS Detection for Resource-Constrained Edge Networks

DGX agent

arXiv:2309.05646v2 Announce Type: replace-cross Abstract: Distributed Denial of Service (DDoS) attacks remain a persistent threat to the availability of Internet services, edge networks, and cyber-phy

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

LinAlg-Bench: A Forensic Benchmark Revealing Structural Failure Modes in LLM Mathematical Reasoning

DGX agent

arXiv:2605.16675v1 Announce Type: new Abstract: We introduce LinAlg-Bench, a diagnostic benchmark evaluating 10 frontier large language models on structured linear algebra computation across a strict

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LiTS: A Modular Framework for LLM Tree Search

DGX agent

arXiv:2603.00631v2 Announce Type: replace Abstract: LiTS is a modular Python framework for LLM reasoning via tree search. It decomposes tree search into three reusable components (Policy, Transition,

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Live from Code with Claude London: we're launching self-hosted sandboxes (public beta) and MCP tunnels (research preview) in Claude Managed …

DGX agent

Live from Code with Claude London: we're launching self-hosted sandboxes (public beta) and MCP tunnels (research preview) in Claude Managed Agents. Run agents inside your own perimeter, with your secu

model-releasesboris-cherny--x
19 May 2026
Model Releases

LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injectio

DGX agent

arXiv:2605.17986v1 Announce Type: cross Abstract: AI agents such as OpenClaw are increasingly deployed in local workflows with access to external tools. This creates indirect prompt-injection (IPI) ri

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

llm-gemini 0.32

DGX agent

llm-gemini 0.32 is an alpha release of Simon Willison's LLM Python library and CLI tool that provides access to Google's Gemini models , continuing work on major architectural changes to support newer

model-releasessimon-willison
19 May 2026
Model Releases

llm-gemini 0.32a0

DGX agent

I don't have current information about this specific entry, so I'll describe what it likely covers based on the available details. This entry documents the release or update of llm-gemini version 0.32

model-releasessimon-willison
19 May 2026
Model Releases

LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models

DGX agent

arXiv:2605.17653v1 Announce Type: cross Abstract: Sub-billion-parameter Transformer language models are increasingly deployed on edge devices, where the privacy, latency, and operating-cost advantages

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations

DGX agent

arXiv:2605.16538v1 Announce Type: cross Abstract: This paper examines the opportunities, limitations, and practical considerations associated with the use of large language models (LLMs) in qualitativ

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

LongMINT: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems

DGX agent

arXiv:2605.18565v1 Announce Type: cross Abstract: Real-world agents operate over long and evolving horizons, where information is repeatedly updated and may interfere across memories, requiring accura

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LoopQ: Quantization for Recursive Transformers

DGX agent

arXiv:2605.16343v1 Announce Type: cross Abstract: Looped language models (LoopLMs) improve parameter efficiency by recursively reusing Transformer blocks, enabling deeper computation under a fixed mod

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

M^2FedAQI: Multimodal Federated Learning for Air Quality Prediction on Heterogeneous Edge Devices

DGX agent

arXiv:2605.16375v1 Announce Type: new Abstract: Accurate air quality prediction is essential for public health, environmental monitoring, and industrial safety. However, most existing approaches rely

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Machine Unlearning for Masked Diffusion Language Models

DGX agent

arXiv:2605.18253v1 Announce Type: cross Abstract: Recent masked diffusion language models (MDLMs), such as LLaDA and Dream, have achieved performance comparable to autoregressive large language models

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MADP: A Multi-Agent Pipeline for Sustainable Document Processing with Human-in-the-Loop

DGX agent

arXiv:2605.17159v1 Announce Type: new Abstract: Document processing automation remains a critical challenge in enterprise environments, where traditional manual approaches are labor-intensive and erro

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics

DGX agent

arXiv:2605.18617v1 Announce Type: cross Abstract: Most existing vision-language manipulation research targets rigid robotic arms, whose fixed morphology limits adaptability in cluttered or confined sp

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MANTA: Multi-turn Assessment for Nonhuman Thinking & Alignment

DGX agent

arXiv:2605.16301v1 Announce Type: cross Abstract: Single-turn benchmarks such as AnimalHarmBench (AHB) have established important baselines for measuring animal welfare alignment in large language mod

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MARS: Technical Report for the CASTLE Challenge at EgoVis 2026

DGX agent

arXiv:2605.18176v1 Announce Type: cross Abstract: This report presents MARS, short for Multimodal Agentic Reasoning with Source selection, our system for the CASTLE Challenge at EgoVis 2026. Participa

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MAVEN A Multi-Agent Framework for Multicultural Text-to-Video Generation

DGX agent

arXiv:2605.16716v1 Announce Type: cross Abstract: Text-to-video (T2V) generation has rapidly progressed in visual fidelity, yet its ability to faithfully represent multiple cultures within a single pr

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MCQ Difficulty Prediction via Modeling Learner Heterogeneity Using Data-Driven Cognitive Profiling

DGX agent

arXiv:2605.16290v1 Announce Type: cross Abstract: Predicting the difficulty of multiple-choice questions (MCQs) is important for effective assessment, yet current methods typically assume a unimodal s

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

DGX agent

arXiv:2603.05308v2 Announce Type: replace-cross Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and claim verification. While large language model

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Membership Inference Attacks on Discrete Diffusion Language Models

DGX agent

arXiv:2605.16445v1 Announce Type: cross Abstract: Masked Diffusion Language Models MDLMs replace autoregressive generation with iterative demasking and their privacy properties are largely unstudied.

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MemOCR: Layout-Aware Visual Memory for Efficient Long-Horizon Reasoning

DGX agent

arXiv:2601.21468v5 Announce Type: replace Abstract: Long-horizon agentic reasoning necessitates effectively compressing growing interaction histories into a limited context window. Most existing memor

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MentalBench: A DSM-Grounded Benchmark for Evaluating Psychiatric Diagnostic Capability of Large Language Models

DGX agent

arXiv:2602.12871v2 Announce Type: replace Abstract: Large language models (LLMs) have attracted growing interest as supportive tools for psychiatric assessment and clinical decision support. However,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting

DGX agent

arXiv:2510.10528v3 Announce Type: replace Abstract: Large reasoning models (LRMs) have demonstrated remarkable proficiency in tackling complex tasks through step-by-step thinking. However, this length

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation

DGX agent

arXiv:2605.17292v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems have shown promise for solving complex tasks through agent collaboration. However, existing frameworks as

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MiniGPT: Rebuilding GPT from First Principles

DGX agent

arXiv:2605.17398v1 Announce Type: new Abstract: This paper presents MiniGPT, a compact from-scratch implementation of GPT-style autoregressive language modeling in PyTorch. The aim is to rebuild the c

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery

DGX agent

arXiv:2605.17198v1 Announce Type: cross Abstract: To be useful for downstream applications, vision decoding models that are trained to reconstruct seen images from human brain activity must be able to

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

DGX agent

arXiv:2601.08118v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as human simulators, both for evaluating conversational systems and for generating fine-tuning da

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Mistral acquires Vienna-based Emmi AI for an undisclosed sum to boost its industrial offerings in Europe; Emmi raised €15M in Austria's largest round in 2025 (Reuters)

DGX agent

Reuters: Mistral acquires Vienna-based Emmi AI for an undisclosed sum to boost its industrial offerings in Europe; Emmi raised €15M in Austria's largest round in 2025 — Europe's leading artificial int

model-releasestechmeme
19 May 2026
Model Releases

MixSD: Mixed Contextual Self-Distillation for Knowledge Injection

DGX agent

arXiv:2605.16865v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to inject new knowledge into language models, but it often degrades pretrained capabilities such as reasonin

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource

DGX agent

arXiv:2506.12119v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) language models dramatically expand model capacity and achieve remarkable performance without increasing per-token co

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Mixture of Experts for Low-Resource LLMs

DGX agent

arXiv:2605.17598v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enable efficient model scaling, yet expert routing behavior across underrepresented languages remains poorly unde

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility

DGX agent

arXiv:2605.16616v1 Announce Type: new Abstract: Autonomous research systems capable of generating complete scientific manuscripts have advanced rapidly, yet robust and realistic evaluation frameworks

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

DGX agent

arXiv:2602.22667v2 Announce Type: replace Abstract: Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abunda

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

MorphSeek: Fine-grained Latent Representation-Level Policy Optimization for Deformable Image Registration

DGX agent

arXiv:2511.17392v3 Announce Type: replace Abstract: Deformable image registration (DIR) remains a fundamental yet challenging problem in medical image analysis, largely due to the prohibitively high-d

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Multi-Party Multi-Objective Optimization as Consensus Search: Runtime Analysis of Cross-Party Recombination

DGX agent

arXiv:2605.17454v1 Announce Type: new Abstract: Multi-party multi-objective optimization problems (MPMOPs) require consensus among autonomous decision makers and therefore differ from flattened many-o

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Multi-site PPG: An In-the-Wild Physiological Dataset from Emerging Multi-site Wearables

DGX agent

arXiv:2605.17859v1 Announce Type: cross Abstract: Wearables are widely used for mobile health monitoring, and photoplethysmography (PPG) is a key sensing modality for heart rate and related physiologi

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Multilingual jailbreaking of LLMs using low-resource languages

DGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models

DGX agent

arXiv:2605.16409v1 Announce Type: cross Abstract: Optical character recognition (OCR) and multilingual text understanding remain major failure modes of multimodal large language models (MLLMs), partic

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

DGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

My notes on Gemini 3.5 Flash - 3x the price of Gemini 3 Flash but Google are planning to use it for many of their own products https://simon…

DGX agent

Google's Gemini 3.5 Flash model costs approximately 3x more than Gemini 3 Flash, despite being a newer version. Google plans to integrate Gemini 3.5 Flash into many of their own products, suggesting t

model-releasessimon-willison--x
19 May 2026
← Previous
1…297298299300301…472
Next →