AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,611 results
Model Releases

ReMMD: Realistic Multilingual Multi-Image Agentic Verification for Multimodal Misinformation Detection

DGX agent

arXiv:2606.24112v1 Announce Type: new Abstract: Multimodal misinformation detection is increasingly important because viral posts now combine long multilingual narratives, several images, mixed proven

model-releasesarxiv-cs-ai
24 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Representation Interventions Enable Lifelong Knowledge Memory Control in LLMs

DGX agent

arXiv:2511.20892v4 Announce Type: replace Abstract: Large language models (LLMs) often produce incorrect or outdated content after being employed. Efficient and accurate knowledge updates without cost

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

RetiSEM: Generalising Causal Models for Fragmented Biomedical Data

DGX agent

arXiv:2606.24488v1 Announce Type: cross Abstract: Learning causal models from fragmented biomedical data is challenging because clinical, molecular, and imaging variables are often incomplete or not j

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients

DGX agent

arXiv:2606.24774v1 Announce Type: new Abstract: Vision-Language Large Models (VLLMs) trained on massive crawled corpora raise pressing copyright and data-provenance concerns. These concerns are partic

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Reward-Centered ReST-MCTS: A Robust Decision-Making Framework for Robotic Manipulation in High Uncertainty Environments

DGX agent

arXiv:2503.05226v2 Announce Type: replace-cross Abstract: Monte Carlo tree search is attractive for robotic manipulation because it can improve action selection through simulation without requiring a

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Right alongside Cursor, Devin Desktop (Windsurf) and CLI now support GLM-5.2 as well. FrontierCode Extended is a benchmark we care deeply ab…

DGX agent

Right alongside Cursor, Devin Desktop (Windsurf) and CLI now support GLM-5.2 as well. FrontierCode Extended is a benchmark we care deeply about for real-world engineering tasks, so it's great to see G

model-releaseszhipu-ai--x
24 Jun 2026
Model Releases

RoPE-Aware Bit Allocation for KV-Cache Quantization

DGX agent

arXiv:2606.24033v1 Announce Type: cross Abstract: Existing low-bit KV-cache quantizers often treat each cached key as a flat vector. Under RoPE, however, a key's contribution to a future attention log

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Rule2Text: A Framework for Generating and Evaluating Natural Language Explanations of Knowledge Graph Rules

DGX agent

arXiv:2508.10971v2 Announce Type: replace-cross Abstract: Knowledge graphs (KGs) can be enhanced through rule mining; however, the resulting logical rules are often difficult for humans to interpret d

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SAFARI: Scaling Long Horizon Agentic Fault Attribution via Active Investigation

DGX agent

arXiv:2606.24626v1 Announce Type: new Abstract: As autonomous agents tackle increasingly complex multi-step, multi-agent tasks, their execution trajectories have scaled beyond the constraints of even

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Sakana AI CEO David Ha (@hardmaru) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical…

DGX agent

Sakana AI CEO David Ha (@hardmaru) appeared on TBS CROSS DIG’s “1on1 Tech.” He talks about our founding story, latest research and technical vision, product launches including Sakana Fugu, Japan’s AI

model-releasesdavid-ha--x
24 Jun 2026
Model Releases

Sample session: https://claude.ai/code/session_01JptTG6Pr3NGWxQXMM3vWvi

DGX agent

This entry references a Claude.ai coding session shared by Simon Willison, a well-known technology writer and open source developer, likely demonstrating a practical example of using Claude's code int

model-releasessimon-willison--x
24 Jun 2026
Model Releases

SciZoom: A Large-scale Benchmark for Hierarchical Scientific Summarization across the LLM Era

DGX agent

arXiv:2603.16131v2 Announce Type: replace Abstract: The explosive growth of AI research has created unprecedented information overload, increasing the demand for scientific summarization at multiple l

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

SEAGAN: domain-Specific and Edge-Aware Graph Attention Network for Dynamic Plant Processes

DGX agent

arXiv:2606.19623v2 Announce Type: replace Abstract: Graph neural networks (GNNs) offer a flexible framework for learning from scientific data with physical, biological, or functional associations. One

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Self-Recognition Finetuning can Prevent and Reverse Emergent Misalignment

DGX agent

arXiv:2606.23700v1 Announce Type: cross Abstract: Emergent misalignment (EM) has been linked to the activation of misaligned persona vectors and evil character traits, suggesting that EM operates thro

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SER: Learning to Ground Video Reasoning with Semantic Evidence Rewards

DGX agent

arXiv:2606.24726v1 Announce Type: new Abstract: Video MLLMs often struggle with fine-grained spatio-temporal reasoning, sometimes generating correct answers based on irrelevant frames or objects. Alth

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

SICI: A Semantic-Pragmatic Complexity Index Reveals Regime Shifts in LLM Stance Detection

DGX agent

arXiv:2606.13189v2 Announce Type: replace Abstract: Prompt-based LLMs are increasingly used for stance detection, but harder examples are not always repaired by clearer instructions, reasoning prompts

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

SignNet-1M: Large-Scale Multilingual Sign Language Video Dataset with Downstream Benchmarks

DGX agent

arXiv:2606.24361v1 Announce Type: new Abstract: Sign language models are typically trained on datasets captured under constrained conditions, with limited viewpoint, background, and signer-identity di

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

simonw/browser-compat-db

DGX agent

simonw/browser-compat-db Inspired by Mozilla's new MDN MCP service - source code here - I decided to try converting their comprehensive mdn/browser-compat-data repository full of browser compatibility

model-releasessimon-willison
24 Jun 2026
Model Releases

Sources: Google AI researchers Jonas Adler and Alexander Pritzel, both viewed internally as key contributors to Gemini, are planning to leave for Anthropic (Bloomberg)

DGX agent

Bloomberg: Sources: Google AI researchers Jonas Adler and Alexander Pritzel, both viewed internally as key contributors to Gemini, are planning to leave for Anthropic — Two leading artificial intellig

model-releasestechmeme
24 Jun 2026
Model Releases

Sources: in a letter to US officials, Anthropic accused Alibaba of adversarial distillation, accessing Claude 28.8M times from April to June via ~25K accounts (Maggie Eastland/Bloomberg)

DGX agent

Maggie Eastland / Bloomberg: Sources: in a letter to US officials, Anthropic accused Alibaba of adversarial distillation, accessing Claude 28.8M times from April to June via ~25K accounts — Anthropic

model-releasestechmeme
24 Jun 2026
Model Releases

SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics Analysis

DGX agent

arXiv:2606.24235v1 Announce Type: new Abstract: Spatial proteomics enables single-cell-resolution characterization of protein expression within tissue architecture, playing a critical role in understa

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Structural Kolmogorov-Arnold Convolutions: Learnable Function on the Values or the Filter Shape as Parameter-Efficient Alternative to Per-Edge Convolutional KANs

DGX agent

arXiv:2606.24371v1 Announce Type: cross Abstract: Convolutional Kolmogorov--Arnold Networks (KANs) replace the fixed weights of a convolutional kernel with learnable univariate functions. The dominant

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SURGELLM: Rethinking Multi-Task Evaluation through Task-Aware Feature Gating with Class-Balanced Normalization

DGX agent

arXiv:2606.24259v1 Announce Type: cross Abstract: Fine-tuned encoders deployed across heterogeneous NLP tasks face three compounding problems: mismatched inductive biases, class-imbalance corruption o

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Synergizing Physically Constrained MCMC and Chemical-Informed Gaussian Processes for Reaction Network Discovery

DGX agent

arXiv:2606.23757v1 Announce Type: cross Abstract: Extracting interpretable governing equations from sparse, noisy chemical time-series data remains difficult because discrete reaction topology and con

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Systematic Exploration of 4-Expert Heterogeneous Mixture-of-Experts via Automated Pipeline Search

DGX agent

arXiv:2606.23739v1 Announce Type: cross Abstract: We present an automated large-scale search pipeline for heterogeneous 4-Expert Mixture-of-Experts (MoE4) architectures within the LEMUR neural network

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

T2D-Bench: Evidence-Gated Evaluation of LLM Outputs for Type 2 Diabetes Using a Multi-Layer Clinical-Lifestyle Knowledge Graph

DGX agent

arXiv:2606.24145v1 Announce Type: new Abstract: Large language models (LLMs) can produce clinically fluent recommendations for type 2 diabetes while failing to satisfy guideline constraints or explici

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Ten Digits on a Train: AI-Assisted Verification of Two Eigenvalue Problems

DGX agent

arXiv:2606.23821v1 Announce Type: cross Abstract: Accurate numerical eigenvalues are often difficult to certify, especially in singular or non-normal settings. This article reports a human--AI collabo

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

The African Language Tax: Quantifying the Cost, Latency, and Context Penalty of Tokenizing African Languages in Frontier LLMs

DGX agent

arXiv:2606.24460v1 Announce Type: cross Abstract: Commercial large language models bill, scale latency, and budget context per token. Yet tokenizers assign more subword tokens to the same meaning in s

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

The AI future is Co:here

DGX agent

The AI future is Co:here on an overheated extremely delayed tgv train in france unable to understand any of the announcements in french and my google translate live translate is absolutely bewildered.

model-releasescohere--x
24 Jun 2026
Model Releases

The Degeneracy Distillery

DGX agent

arXiv:2606.23838v1 Announce Type: new Abstract: When two or more parameters or labels produce similar data, they are degenerate, or hard to distinguish. Degeneracies render both label prediction and i

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

The Kimi API is now live on AWS Marketplace. 🚀 If your team is already running on AWS, you can now access Kimi with consolidated billing. P…

DGX agent

The Kimi API is now live on AWS Marketplace. 🚀 If your team is already running on AWS, you can now access Kimi with consolidated billing. Plus, eligible customers can apply Kimi API usage directly tow

model-releaseskimi-moonshot--x
24 Jun 2026
Model Releases

This is the strongest ARC-AGI-2 performance to date by an open-source model.

DGX agent

This is the strongest ARC-AGI-2 performance to date by an open-source model. GLM-5.2 from @Zai_org on ARC-AGI (Verified) - ARC-AGI-2: 22.8%, 0.25 - ARC-AGI-1: 77.0%, 0.19 Performance is comparable wit

model-releasesfrancois-chollet--x
24 Jun 2026
Model Releases

TIGER: Taming Identity, Geometry, and Generative Priors for High-Quality Face Video Restoration

DGX agent

arXiv:2606.24336v1 Announce Type: new Abstract: Face Video Restoration (FVR) aims to recover high-fidelity facial videos from degraded input while preserving identity and semantic consistency across f

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias

DGX agent

arXiv:2606.24596v1 Announce Type: new Abstract: As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling

DGX agent

arXiv:2606.24187v1 Announce Type: new Abstract: Long video understanding remains a daunting challenge for Multimodal Large Language Models (MLLMs) due to the excessive computation and memory footprint

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Towards Spec Learning: Inference-Time Alignment from Preference Pairs

DGX agent

arXiv:2606.24004v1 Announce Type: cross Abstract: Steering a large language model (LLM) toward a desired behavior typically relies on an iterative process of hand-crafting a prompt based on a careful

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Transformer-Based Language Models Across Domain Verticals: Architectures, Applications and Critical Assessment

DGX agent

arXiv:2606.24331v1 Announce Type: new Abstract: Transformer-based language models have become the default substrate for natural language processing and the pace of new releases has made it hard for pr

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Tri-Efficient Transfer Learning for Point Cloud Videos

DGX agent

arXiv:2606.24175v1 Announce Type: new Abstract: While point cloud foundation models have significantly advanced point cloud video understanding, existing parameter-efficient fine-tuning (PEFT) methods

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Trimming the Long-Tail of Visual World Modeling Evaluation

DGX agent

arXiv:2606.24256v1 Announce Type: new Abstract: Physical interactions follow a long-tailed distribution: a set of common and regular interactions dominates human experience and visual data, while a br

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

TrOCR for Medieval HTR: A Systematic Ablation Study with Cross-Dataset Validation

DGX agent

arXiv:2606.24302v1 Announce Type: new Abstract: Fine-tuning transformer-based handwritten text recognition (HTR) models on medieval manuscripts is challenging because these models are pre-trained on m

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Try Kimi K2.7 and GLM 5.2 for free in Devin Desktop and CLI

DGX agent

Try Kimi K2.7 and GLM 5.2 for free in Devin Desktop and CLI Kimi K2.7 Code and GLM 5.2 are available in Devin Desktop and CLI Both perform strongly on FrontierCode Extended, our benchmark for real-wor

model-releasescognition-ai--x
24 Jun 2026
Model Releases

Tuning without Peeking: Provable Generalization Bounds and Robust LLM Post-Training

DGX agent

arXiv:2507.01752v4 Announce Type: replace-cross Abstract: Gradient-based optimization is the workhorse of deep learning, offering efficient and scalable training via backpropagation. However, exposing

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving

DGX agent

arXiv:2606.24759v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong potential for autonomous driving scene understanding, yet existing methods still fac

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

UniRED: Unified RGB-D Video Frame Interpolation with Event Guidance

DGX agent

arXiv:2606.24282v1 Announce Type: new Abstract: High frame-rate RGB-D videos are crucial for a variety of downstream tasks, including motion analysis, dynamic scene understanding, and 3D reconstructio

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Upbound open-sources Modelplane to optimize inference clusters

DGX agent

Upbound Inc. today released Modelplane, a new open-source tool for managing artificial intelligence inference clusters. San Francisco-based Upbound is backed by $69 million from Alphabet Inc.’s GV fun

model-releasessiliconangle
24 Jun 2026
Model Releases

Variational Model Merging for Pareto Front Estimation in Multitask Finetuning

DGX agent

arXiv:2412.08147v2 Announce Type: replace-cross Abstract: Pareto fronts are useful to find good task-mixing strategies for multitask finetuning, but they are also costly to compute. To reduce costs, r

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

VeriPilot: An LLM-Powered Verilog Debugging Framework

DGX agent

arXiv:2606.23759v1 Announce Type: cross Abstract: Verilog debugging remains one of the most time-consuming stages in digital circuit design. Recent advances in Large Language Models (LLMs) have enable

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

video-SALMONN-R^3: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding

DGX agent

arXiv:2606.24477v1 Announce Type: cross Abstract: Video large language models (LLMs) are often constrained by computation and memory budgets, leading them to use reduced frame rates and spatial resolu

model-releasesarxiv-cs-ai
24 Jun 2026
← Previous
1…173174175176177…472
Next →