AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

How Far Can Disaggregation Go? A Design-Space Exploration of Attention-FFN Disaggregation for Efficient MoE LLM Serving

DGX agent

arXiv:2605.28302v1 Announce Type: cross Abstract: Modern large language model (LLM) inference has progressively disaggregated to keep pace with growing model sizes and tight TTFT and TPOT service-leve

model-releasesarxiv-cs-ai
28 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

how it started how it’s going

DGX agent

'How it started, how it's going' is a popular internet meme format that compares two contrasting images or states—typically showing an initial hopeful or humble beginning alongside a current outcome t

model-releasescohere--x
28 May 2026
Model Releases

How the University of Central Oklahoma is using AI to streamline analysis of complex criminal cases

DGX agent

In the high-stakes world of forensic science, time is the enemy of justice. The University of Central Oklahoma (UCO) Forensic Science Institute (FSI) was looking for an innovative AI solution that cou

model-releasesgoogle-cloud-ai
28 May 2026
Model Releases

https://github.com/run-llama/liteparse これか。日本語PDFでどんなか試しとこう。

DGX agent

https://github.com/run-llama/liteparse これか。日本語PDFでどんなか試しとこう。 We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf

model-releasesjerry-liu--x
28 May 2026
Model Releases

https://mistral.ai/news/ai-now-summit-2026/

DGX agent

Mistral AI announced its participation in or perspective on the AI Now Summit 2026, likely discussing developments in AI safety, ethics, or industry trends relevant to the conference. The announcement

model-releasesmistral-ai--x
28 May 2026
Model Releases

HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning

DGX agent

arXiv:2605.27724v1 Announce Type: cross Abstract: Imitation learning is a promising approach for training humanoid robots to both walk and manipulate, but it requires a large number of demonstrations,

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Hurwitz Quaternion Multiplicative Quantization for KV Cache Compression

DGX agent

arXiv:2605.27646v1 Announce Type: cross Abstract: We propose extbf{Hurwitz Quaternion Multiplicative Quantization (HQMQ)}, a extbf{calibration-free} method for KV cache compression of large language m

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

I had early access to Opus 4.8. Was impressed by it. Here is Opus 4.8's one shot of 'create a visually interesting shader that can run in tw…

DGX agent

I had early access to Opus 4.8. Was impressed by it. Here is Opus 4.8's one shot of 'create a visually interesting shader that can run in twigl, make it like an infinite city of neo-gothic towers part

model-releasessonya-huang--x
28 May 2026
Model Releases

I had Opus 4.8 in Claude Code write a sophisticated, if minor, academic paper from a archive of hundreds of de-identified research files fro…

DGX agent

I had Opus 4.8 in Claude Code write a sophisticated, if minor, academic paper from a archive of hundreds of de-identified research files from years ago I had to use GPT-5.5 Pro as a reviewer, it spott

model-releasesethan-mollick--x
28 May 2026
Model Releases

I signed up for another SaaS

DGX agent

Ben Kamarov reflects on his decision to sign up for yet another SaaS product, likely discussing his evaluation criteria, the specific tool's features, or lessons learned about SaaS adoption and tool p

model-releasesben-s-bites
28 May 2026
Model Releases

I think you’ll really like Opus 4.8 It’s as smart as its benchmarks show but expresses and utilizes that intelligence in a warm and collabor…

DGX agent

I think you’ll really like Opus 4.8 It’s as smart as its benchmarks show but expresses and utilizes that intelligence in a warm and collaborative way. Workflows are a great way to utilize it- I’m hook

model-releasesthariq--x
28 May 2026
Model Releases

I tried the liteparse's web browser version today to convert a couple of PDF to text and was shocked at the speed. I had to recheck twice to…

DGX agent

I tried the liteparse's web browser version today to convert a couple of PDF to text and was shocked at the speed. I had to recheck twice to see whether it even did the complete processing or not 😅 ht

model-releasesjerry-liu--x
28 May 2026
Model Releases

IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction

DGX agent

arXiv:2510.06928v2 Announce Type: replace Abstract: Autoregressive models have emerged as a powerful paradigm for visual content creation, but often overlook the intrinsic structural properties of vis

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

IBM and Red Hat commit $5B to establish a new model for open-source software, dubbed Project Lightwell, and will deploy 20,000 engineers, supported by AI (Connor Hart/Wall Street Journal)

DGX agent

Connor Hart / Wall Street Journal: IBM and Red Hat commit $5B to establish a new model for open-source software, dubbed Project Lightwell, and will deploy 20,000 engineers, supported by AI — Project L

model-releasestechmeme
28 May 2026
Model Releases

if you replace billions with millions, this sounds like any other high-growth startup fundraise announcement 😉

DGX agent

if you replace billions with millions, this sounds like any other high-growth startup fundraise announcement 😉 We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by

model-releasesjerry-liu--x
28 May 2026
Model Releases

IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following

DGX agent

arXiv:2605.28218v1 Announce Type: new Abstract: Modern translation workflows demand more than semantic equivalence. Users routinely require models to preserve JSON or HTML schemas, honor curated gloss

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This …

DGX agent

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This is an exciting milestone for Glean, and it's a signal about wh

model-releasessonya-huang--x
28 May 2026
Model Releases

Imitation Learning for Robot Assistance in Open Surgery: A Multi-Policy Evaluation on Suture Following

DGX agent

arXiv:2605.28736v1 Announce Type: new Abstract: This study presents the first evaluation of general-purpose imitation learning for surgeon-robot collaborative assistance in open surgery, targeting sut

model-releasesarxiv-cs-ro
28 May 2026
Model Releases

In the last 30 days alone: – Microsoft cancelled most of its Claude Code licenses, citing cost – Uber burned through its entire 2026 AI budg…

DGX agent

In the last 30 days alone: – Microsoft cancelled most of its Claude Code licenses, citing cost – Uber burned through its entire 2026 AI budget in 4 months – Uber's COO publicly said AI costs are 'hard

model-releasesgary-marcus--x
28 May 2026
Model Releases

Inpainting-Style Conditional Diffusion for Multivariable Time Series Forecasting

DGX agent

arXiv:2605.28324v1 Announce Type: new Abstract: In this paper, we propose a novel conditional diffusion-based framework for multivariable time-series solar power forecasting. The proposed method refor

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Integrated and Cross-Architecture Interpretation of LLM Reasoning

DGX agent

arXiv:2605.28006v1 Announce Type: cross Abstract: Understanding how LLMs reason is hindered by a practical asymmetry: while their generated outputs are observable, the underlying reasoning patterns re

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Internally Referenced Low-Light Enhancement

DGX agent

arXiv:2605.28605v1 Announce Type: new Abstract: Self-supervised low-light image enhancement (LLIE) is highly appealing as it eliminates the reliance on external paired data. However, the lack of exter

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Interpretability-Guided Layer Selection over Subspace Projection: SAEs as Stethoscopes, Not Scalpels, for Raw Task Vector Model Editing

DGX agent

arXiv:2605.28649v1 Announce Type: cross Abstract: LLMs increasingly require surgical model editing to enhance domain-specific capabilities without incurring the computational cost or catastrophic forg

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

IPO-Mine: A Toolkit and Dataset for Section-Structured Analysis of Long, Multimodal IPO Documents

DGX agent

arXiv:2605.28714v1 Announce Type: cross Abstract: An Initial Public Offering (IPO) filing is a document released when a private firm goes public, allowing individual (retail) investors to purchase its

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage

DGX agent

arXiv:2605.28247v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a key technique for en- hancing LLM reasoning, yet its data ineffi- ciency remains a

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Janus-LoRA: A Balanced Low-Rank Adaptation for Continual Learning

DGX agent

arXiv:2605.28495v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has emerged as a promising paradigm for Continual Learning. It independently updates its low-rank factors (A and B), creating

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

JMedEthicBench: A Multi-Turn Conversational Benchmark for Evaluating Medical Safety in Japanese Large Language Models

DGX agent

arXiv:2601.01627v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare field, it becomes essential to carefully evaluate their medical safety

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datas…

DGX agent

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datasets, evals and autoresearch runs as we can in our pursuit of

model-releasesclem-delangue--x
28 May 2026
Model Releases

KSAFE-MM: A Multimodal Safety Benchmark via Localized Contextualization for Korean Cultural Risks

DGX agent

arXiv:2605.28013v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exacerbate safety risks by introducing vulnerabilities across multiple modalities, such as language and vision.

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs

DGX agent

arXiv:2605.27984v1 Announce Type: cross Abstract: Speech language models (SpeechLMs) have achieved substantial progress by extending large language models (LLMs) to the speech modality. However, Speec

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Laguna M.1/XS.2 Technical Report

DGX agent

arXiv:2605.27605v1 Announce Type: new Abstract: We present Laguna M.1 and Laguna XS.2, two Mixture-of-Experts foundation models built for long-horizon, agentic coding: M.1 has 225.8B total parameters

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Law of Neural Interaction: Depth-Width Shape, Interaction Efficiency, and Generalization

DGX agent

arXiv:2605.27989v1 Announce Type: new Abstract: The guidance of scaling laws has increased the resource demands of modern large language models (LLMs), yet it remains questionable whether these models

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

LCO: LLM-based Constraint Optimization for Safer Agentic LLMs in Real-world Tasks

DGX agent

arXiv:2605.27375v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly acting as autonomous agents, but their continuous interaction with the environment can lead to in-context

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Learning Compositional Latent Structure with Vector Networks

DGX agent

arXiv:2605.28007v1 Announce Type: cross Abstract: Deep networks are powerful function approximators, but they typically store many different computations in shared weight matrices, making it difficult

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Learning to Translate from Soft to Hard LLM Prompts

DGX agent

arXiv:2605.27642v1 Announce Type: new Abstract: Soft prompt tuning is a parameter-efficient method for adapting LLMs to specific tasks, but suffers from a lack of interpretability. Building on recent

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

LEIA: Learned Environment for Interactive Architected Materials

DGX agent

arXiv:2605.28368v1 Announce Type: new Abstract: World models have enabled interactive exploration of game environments and robotic manipulation, but physical engineering remains beyond their reach: re

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

LESA: Learnable Stage-Aware Predictors for Diffusion Model Acceleration

DGX agent

arXiv:2602.20497v3 Announce Type: replace-cross Abstract: Diffusion models have achieved remarkable success in image and video generation tasks. However, the high computational demands of Diffusion Tr

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Let the Results Speak: A Replication-First Paradigm for LLM Behavioral Benchmarking

DGX agent

arXiv:2605.27914v1 Announce Type: cross Abstract: Subjective evaluation of LLM behavior -- empathy, restraint, calibrated emotional tone -- is hard. Human inter-rater agreement on such qualities satur

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

LiveBrowseComp: Are Search Agents Searching, or Just Verifying What They Already Know?

DGX agent

arXiv:2605.28721v1 Announce Type: new Abstract: Are LLM-based search agents genuinely searching, or using the web to verify what they already know? We study this question on BrowseComp with three diag

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

llm-anthropic 0.25.1

DGX agent

Release: llm-anthropic 0.25.1 New model: Claude Opus 4.8 (claude-opus-4.8). New -o fast 1 option for fast mode, for organizations with that feature enabled on their account. Default max_tokens for eac

model-releasessimon-willison
28 May 2026
Model Releases

LLM Zeroth-Order Fine-Tuning is an Inference Workload

DGX agent

arXiv:2605.28760v1 Announce Type: new Abstract: Zeroth-order (ZO) fine-tuning is attractive for large language models because it replaces backpropagation with forward objective evaluations. Existing i

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Local MDI+: Local Feature Importances for Tree-Based Models

DGX agent

arXiv:2506.08928v2 Announce Type: replace Abstract: Tree-based ensembles such as random forests remain the go-to for tabular data over deep learning models due to their prediction performance and comp

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Tha…

DGX agent

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Thanks @andimarafioti for the blog post on how to set this up: h

model-releasesclem-delangue--x
28 May 2026
Model Releases

LV-OSD: Language-Vision-Complementary Open-Set Object Detection

DGX agent

arXiv:2605.28271v1 Announce Type: new Abstract: Object detection is an important task in computer vision, which aims to detect the objects of interest. through the given category list or query images.

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

MACReD: A Multi-Agent Collaborative Reasoning Framework for Reaction Diagram Parsing

DGX agent

arXiv:2605.28077v1 Announce Type: new Abstract: Parsing chemical reaction diagrams from scientific literature is challenging due to heterogeneous layouts, intertwined visual elements, and the difficul

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Mags-RL: Wearing Multimodal LLMs a Magnifying Glass via Agentic Reinforcement Learning For Complex Scene Reasoning

DGX agent

arXiv:2605.27960v1 Announce Type: new Abstract: Despite their popularity and success, Multimodal Large Language Models (MLLMs) often struggle to interpret images accurately, which limits their reasoni

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Mahalanobis PatchCore: Covariance-Aware and Streaming-Compatible Industrial Anomaly Detection

DGX agent

arXiv:2605.27748v1 Announce Type: cross Abstract: Industrial visual anomaly detection is usually one-class: normal images are abundant, while defects are rare, heterogeneous, and often unavailable dur

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation

DGX agent

arXiv:2605.28173v1 Announce Type: new Abstract: End-to-end manga generation is a structured visual storytelling task that requires story decomposition, recurring character and scene grounding, page la

model-releasesarxiv-cs-cv
28 May 2026
← Previous
1…254255256257258…475
Next →