AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

Token-Level LLM Collaboration via FusionRoute

DGX agent

arXiv:2601.05106v4 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strengths across diverse domains. However, achieving strong performance across these domains with a singl

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Tokenization with Split Trees

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.22705v1 Announce Type: new Abstract: We introduce Tokenization with Split Trees (ToaST), a subword tokenization method that directly optimizes compression under a new recursive inference pr

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Towards Clinically Interpretable Ophthalmic VQA via Spatially-Grounded Lesion Evidence

DGX agent

arXiv:2605.22414v1 Announce Type: new Abstract: Visual Question Answering (VQA) holds great promise for clinical support, particularly in ophthalmology, where retinal fundus photography is essential f

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Towards Selection of Large Multimodal Models as Engines for Burned-in Protected Health Information Detection in Medical Images

DGX agent

arXiv:2511.02014v2 Announce Type: replace Abstract: The detection of Protected Health Information (PHI) in medical imaging is critical for safeguarding patient privacy and ensuring compliance with reg

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Training-Trajectory-Aware Token Selection

DGX agent

arXiv:2601.10348v2 Announce Type: replace Abstract: Efficient distillation is a key pathway for converting expensive reasoning capability into deployable efficiency, yet in the frontier regime where t

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Transcription and Recognition of Italian Parliamentary Speeches Using Vision-Language Models

DGX agent

arXiv:2603.28103v2 Announce Type: replace-cross Abstract: Parliamentary proceedings represent a rich yet challenging resource for computational analysis, particularly when preserved only as scanned hi

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation

DGX agent

arXiv:2605.22355v1 Announce Type: new Abstract: Public transit route planning traditionally depends on structured map infrastructure and complex routing engines, and no existing dataset supports train

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Transporting Task Vectors across Different Architectures without Training

DGX agent

arXiv:2602.12952v2 Announce Type: replace-cross Abstract: Adapting large pre-trained models to downstream tasks often produces task-specific parameter updates that are expensive to relearn for every m

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework

DGX agent

arXiv:2605.22620v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervis

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Ultra-High-Definition Image Quality Assessment via Graph Representation Learning

DGX agent

arXiv:2605.22192v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) for ultrahighdefinition (UHD) images remains challenging because native-resolution inference is computationally ex

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Understanding Data Temporality Impact on Large Language Models Pre-training

DGX agent

arXiv:2605.22769v1 Announce Type: new Abstract: Large language models (LLMs) are typically trained on shuffled corpora, yielding models whose knowledge is frozen at train time and whose temporal groun

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

UniVL: Unified Vision-Language Embedding for Spatially Grounded Contextual Image Generation

DGX agent

arXiv:2605.21611v1 Announce Type: new Abstract: We introduce spatially grounded contextual image generation, a controllable image generation task that reframes the conditioning paradigm. Instead of su

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

DGX agent

arXiv:2602.00122v2 Announce Type: replace Abstract: In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive mann

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

VEELA: A Clinically-Constrained Benchmark for Liver Vessel Segmentation in Computed Tomography Angiography

DGX agent

arXiv:2605.22357v1 Announce Type: new Abstract: Accurate segmentation of hepatic and portal vessels in contrast-enhanced computed tomography angiography (CTA) remains challenging due to complex vascul

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis

DGX agent

arXiv:2605.22570v1 Announce Type: new Abstract: Spatio-temporal reasoning is a core capability for Multimodal Large Language Models (MLLMs) operating in the real world. As such, evaluating it precisel

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction

DGX agent

arXiv:2602.13294v3 Announce Type: replace Abstract: Evaluating whether Multimodal Large Language Models (MLLMs) genuinely reason about physical dynamics remains challenging. Most existing benchmarks r

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Visual-Advantage On-Policy Distillation for Vision-Language Models

DGX agent

arXiv:2605.21924v1 Announce Type: new Abstract: On-policy knowledge distillation has proven effective for language models, yet its application to vision-language models (VLMs) remains underexplored. W

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

We are making our discount permanent! 🎉 Enjoy building with DeepSeek-V4-Pro and bring your innovative ideas to life! 🚀

DGX agent

DeepSeek has announced a permanent discount for its DeepSeek-V4-Pro model, encouraging developers to build and innovate with the platform. The announcement was made via social media and emphasizes the

model-releasesjeremy-howard--x
22 May 2026
Model Releases

We desperately need better ways of evaluating models. Something that shows how helpful they are at working hand-in-hand with humans to help …

DGX agent

We desperately need better ways of evaluating models. Something that shows how helpful they are at working hand-in-hand with humans to help them get stuff done in a cooperative/iterative way. The Clau

model-releasesjeremy-howard--x
22 May 2026
Model Releases

We’re taking suggestions on what you want to see next week ✍️

DGX agent

OpenAI solicited community feedback on X regarding content or features they should prioritize in the following week. This post reflects OpenAI's practice of engaging their audience to guide product de

model-releasesopenai--x
22 May 2026
Model Releases

When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering

DGX agent

arXiv:2605.21807v1 Announce Type: new Abstract: Across medical specialties, clinical practice is anchored in evidence-based guidelines that codify best studied diagnostic and treatment pathways. These

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Wordle 1,797 4/6 ⬛⬛🟨🟨🟩 ⬛⬛⬛🟨⬛ 🟨🟨🟨⬛🟩 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result (puzzle #1,797) where the player achieved a solution in 4 attempts, using color-coded emoji feedback (⬛ for incorrect letters, 🟨 for correct letters in wrong p

model-releasesanthropic--x
22 May 2026
Model Releases

X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation

DGX agent

arXiv:2605.21699v1 Announce Type: cross Abstract: Cross-tokenizer knowledge distillation allows a student model to learn from teachers with incompatible vocabularies. Prior work operates on hidden sta

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Yesterday we released Aleph 2.0, our upgraded video editing model that lets you change exactly what you want while keeping everything else t…

DGX agent

Yesterday we released Aleph 2.0, our upgraded video editing model that lets you change exactly what you want while keeping everything else the same. Available inside our new Edit Studio, you can work

model-releasescristobal-valenzuela--x
22 May 2026
Model Releases

A Deployment Audit of Release-Side Risk in Conformal Triage under Prevalence Shift

DGX agent

arXiv:2605.20956v1 Announce Type: new Abstract: Conformal triage converts predictive scores into deployment actions that either release a case, flag it for urgent attention, or defer it to human revie

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

A Free Lunch in LLM Compression: Revisiting Retraining after Pruning

DGX agent

arXiv:2510.14444v3 Announce Type: replace Abstract: Post-training pruning can substantially reduce LLM inference costs, but it often degrades quality unless the remaining weights are adapted. Since gl

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

A strongly annotated passive acoustic dataset for tropical bird monitoring

DGX agent

arXiv:2605.20578v1 Announce Type: cross Abstract: Passive acoustic monitoring enables continuous, non-invasive biodiversity assessment across diverse ecosystems. The scale of these datasets has driven

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

A Unified Framework for Uncertainty-Aware Explainable Artificial Intelligence: A Case Study in Power Quality Disturbance Classification

DGX agent

arXiv:2605.21114v1 Announce Type: new Abstract: Post-hoc explainable AI (XAI) methods typically produce deterministic attribution maps, whereas Bayesian neural networks (BNNs) induce a distribution ov

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

ACL-Verbatim: hallucination-free question answering for research

DGX agent

arXiv:2605.21102v1 Announce Type: new Abstract: Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Ada2MS: A Hybrid Optimization Algorithm Based on Exponential Mixing of Elementwise and Global Second-Moment Estimates

DGX agent

arXiv:2605.20533v1 Announce Type: new Abstract: Optimization algorithms are core methods by which machine learning models iteratively minimize loss functions, update parameters, learn from data, and i

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Adobe, Canva, and CapCut announce Gemini integrations to let users access the companies' image and video editing tools within the Gemini app (James Peckham/PCMag)

DGX agent

James Peckham / PCMag: Adobe, Canva, and CapCut announce Gemini integrations to let users access the companies' image and video editing tools within the Gemini app — Adobe, Canva, and CapCut all plan

model-releasestechmeme
21 May 2026
Model Releases

Adversarial Robustness in One-Stage Learning-to-Defer

DGX agent

arXiv:2510.10988v2 Announce Type: replace-cross Abstract: Learning-to-Defer (L2D) enables hybrid decision-making by routing inputs either to a predictor or to external experts. While promising, L2D is

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

AgentAtlas: Beyond Outcome Leaderboards for LLM Agents

DGX agent

arXiv:2605.20530v1 Announce Type: cross Abstract: Large language model agents now act on codebases, browsers, operating systems, calendars, files, and tool ecosystems, but the benchmarks used to evalu

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control

DGX agent

arXiv:2512.23292v3 Announce Type: replace-cross Abstract: The prevailing paradigm in AI for physical systems (scaling general-purpose foundation models toward universal multimodal reasoning) confronts

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

AGPO: Adaptive Group Policy Optimization with Dual Statistical Feedback

DGX agent

arXiv:2605.20722v1 Announce Type: new Abstract: Reinforcement learning improves LLM reasoning, but PPO/GRPO typically use fixed clipping and decoding temperature, which makes training brittle and tuni

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

AI-Assisted Scientific Assessment: A Case Study on Climate Change

DGX agent

arXiv:2602.09723v2 Announce Type: replace Abstract: The emerging paradigm of AI co-scientists focuses on tasks characterized by repeatable verification, where agents explore search spaces in 'guess an

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

AMAR: Lightweight Attention-Based Multi-User Activity Recognition from Wi-Fi CSI

DGX agent

arXiv:2605.20649v1 Announce Type: cross Abstract: Wi-Fi-based human activity recognition (HAR) has emerged as a promising approach for contactless sensing, leveraging channel state information (CSI) c

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

An exponential mechanism based on quadratic approximations for fine-tuning machine learning models with privacy guarantees

DGX agent

arXiv:2605.20521v1 Announce Type: new Abstract: Fine-tuning adapts a pretrained machine learning model to a small, sensitive dataset, but this process risks memorizing individual new data points, maki

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Anatomy of Agentic Memory: Taxonomy and Empirical Analysis of Evaluation and System Limitations

DGX agent

arXiv:2602.19320v2 Announce Type: replace Abstract: Agentic memory systems enable large language model (LLM) agents to maintain state across long interactions, supporting long-horizon reasoning and pe

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

AnimeAdapter: Fine-grained and Consistent Zero-shot Anime Character Generation

DGX agent

arXiv:2605.20237v1 Announce Type: new Abstract: We present a lightweight appearance adapter for Stable Diffusion that enables controllable and consistent anime character generation under diverse editi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Another win for open-source robotics! 🔥 @huggingface just released a fully open-source humanoid robot, and you can build one for $2,500. I'…

DGX agent

Another win for open-source robotics! 🔥 @huggingface just released a fully open-source humanoid robot, and you can build one for $2,500. I'm a huge advocate of open-source in robotics space. Why? Robo

model-releasesclem-delangue--x
21 May 2026
Model Releases

Anthropic’s Code with Claude showed off coding’s future—whether you like it or not

DGX agent

The vibes were strong at Code with Claude, Anthropic’s two-day event for software developers in London that kicked off on May 19, the same day as Google’s I/O in Palo Alto. (A coincidence, not a flex,

model-releasesmit-tech-review
21 May 2026
Model Releases

APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents

DGX agent

arXiv:2605.21240v1 Announce Type: new Abstract: LLM agents have shown strong performance across a wide range of complex tasks, including interactive environments that require long-horizon decision mak

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

API Keys Are Open Secrets

DGX agent

Today, AI services rely heavily on API keys. To run AI agents, users provide API keys that signify paid tokens, subscriptions, or paid accounts. While API keys are easy to use, it is just as easy to u

model-releasesgoogle-cloud-ai
21 May 2026
Model Releases

APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings

DGX agent

arXiv:2605.21063v1 Announce Type: new Abstract: Typical LLM responses tend to follow a default style, even though users often have distinct preferences regarding tone, verbosity, and formality that th

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Approximation Theory for Neural Networks: Old and New

DGX agent

arXiv:2605.21451v1 Announce Type: new Abstract: Universal approximation theorems provide a mathematical explanation for the expressive power of neural networks. They assert that, under mild conditions

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

ArchSIBench: Benchmarking the Architectural Spatial Intelligence of Vision-Language Models

DGX agent

arXiv:2605.20837v1 Announce Type: new Abstract: Architectural spatial intelligence, the ability to recognize and infer architectural space, is fundamental to tasks such as robot navigation, embodied i

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

AttriStory: Fine-grained Attribute Realization for Visual Storytelling with Diffusion Models

DGX agent

arXiv:2605.20777v1 Announce Type: new Abstract: Visual storytelling with diffusion models has made impressive strides in maintaining character consistency across narrative scenes. However, a critical

model-releasesarxiv-cs-cv
21 May 2026
← Previous
1…279280281282283…472
Next →