AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,767 results
Model Releases

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Ex…

DGX agent

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Expert models are kind of a nice sweet spot and can solve chal

model-releasessebastian-raschka--x
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training

DGX agent

arXiv:2606.26102v1 Announce Type: cross Abstract: Standard post-training pipelines apply supervised fine-tuning (SFT) and reinforcement learning (RL) to make language models helpful, but these process

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Here is Google Gemini talking about Roger Ebert's review of the 2016 film The Jungle Book. Ebert died in 2013. @GaryMarcus

DGX agent

This post highlights an apparent error where Google's Gemini AI attributed a film review to Roger Ebert for a 2016 movie, despite Ebert's death in 2013, making such a review impossible. The post, shar

model-releasesgary-marcus--x
26 Jun 2026
Model Releases

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model c…

DGX agent

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model cheated more than any public model they've tested, and even r

model-releasesdair-ai--x
26 Jun 2026
Model Releases

hisao{}: A GPU-Native Parallel Optimizer for Multimodal Black-Box Functions via Convergence-Anticonvergence Oscillation

DGX agent

arXiv:2606.26164v1 Announce Type: new Abstract: Finding all modes of a multimodal black-box function is a fundamental challenge in optimization, Bayesian inference, and scientific computing. Existing

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

HOB: A Holistically Optimized Bidding Strategy under Heterogeneous Bidding Environments

DGX agent

arXiv:2510.15238v2 Announce Type: replace-cross Abstract: Optimizing a single advertising campaign across heterogeneous channels is a central challenge in industrial autobidding. Auction mechanisms va

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?

DGX agent

arXiv:2606.26346v1 Announce Type: new Abstract: Agentic benchmarks have emerged across general-purpose and domain-specific settings, including finance, coding, law, and drug discovery, yet energy-doma

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Hybrid privacy-aware semantic search: SVD-truncated document geometry and CKKS-encrypted query reranking under a restricted threat model

DGX agent

arXiv:2606.26373v1 Announce Type: cross Abstract: Dense embeddings power semantic search and retrieval-augmented generation, but embedding-inversion attacks can reconstruct source text from a vector:

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

HyperDFlash: MHC-Aligned Block Speculative Decoding with Gated Residual Reduction

DGX agent

arXiv:2606.26744v1 Announce Type: cross Abstract: We present HyperDFlash, a block-parallel speculative decoding framework tailored to the novel multi-hyper-connection (MHC) architecture proposed by De

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

I can personally attest: OpenClaude using GLM 5.2 is now performing on par with Claude Code powered by Opus 4.8.

DGX agent

I cannot verify the claims in this post as the URL format appears invalid and the specific version numbers (GLM 5.2, Claude Code/Opus 4.8) don't correspond to publicly documented model releases as of

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of peop…

DGX agent

I feel like on X all you hear about is elaborate plans by firms to build their own AI stacks but in my experience companies are full of people who want access to Claude or ChatGPT and are pressuring t

model-releasesethan-mollick--x
26 Jun 2026
Model Releases

If your benchmark relies on a static dataset or sampling from a static distribution densely known at training time, then it is fundamentally…

DGX agent

If your benchmark relies on a static dataset or sampling from a static distribution densely known at training time, then it is fundamentally measuring memorization/retrieval. Which might be fine if yo

model-releasesfrancois-chollet--x
26 Jun 2026
Model Releases

Implementation of reinforcement learning in chemical reaction networks: application to phototaxis as curiosity-driven exploration

DGX agent

arXiv:2606.26168v1 Announce Type: new Abstract: Living systems navigate environments using noisy and incomplete sensory signals. In unicellular algae, phototaxis is often modeled as a mechanistic run-

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Information-Aware KV Cache Compression for Long Reasoning

DGX agent

arXiv:2606.26875v1 Announce Type: cross Abstract: Reasoning capability has advanced rapidly in large language models (LLMs), leading to an increasing size of key-value (KV) cache in both prefilling an

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Inherited Circuits, Learned Semantics: How Fine-Tuning Creates Evasion Vulnerabilities Invisible to Standard Evaluation

DGX agent

arXiv:2606.27091v1 Announce Type: cross Abstract: LLMs fine-tuned for security classification are usually evaluated on held-out examples from the same distribution as their training data. We show that

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Instruction Bleed: Cross-Module Interference in Prompt-Composed Agentic Systems

DGX agent

arXiv:2606.26356v1 Announce Type: new Abstract: Practitioners of prompt-composed agentic systems report a recurring failure mode: editing one prompt module silently shifts the behavior of others despi

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Introducing a limited preview of GPT-5.6 Sol, our next generation frontier model, as well as GPT-5.6 Terra, a balanced model for efficient, …

DGX agent

Introducing a limited preview of GPT-5.6 Sol, our next generation frontier model, as well as GPT-5.6 Terra, a balanced model for efficient, everyday work, and GPT-5.6 Luna, a fast and affordable model

model-releasesopenai--x
26 Jun 2026
Model Releases

Jailbreaking for the Average Jane: Choosing Optimal Jailbreaks via Bandit Algorithms for Automatically Enhanced Queries

DGX agent

arXiv:2606.26936v1 Announce Type: cross Abstract: With a profusion of jailbreaks for LLMs now widely known, a growing concern is that non-expert malicious actors ('the average Jane') could elicit acti

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

June Launches + Live Q&A Hear from Comfy's CEO @yoland_yan and product leaders Deep Mehta, Alexis Rolland, @jojodecayz , and Matt Miller who…

DGX agent

June Launches + Live Q&A Hear from Comfy's CEO @yoland_yan and product leaders Deep Mehta, Alexis Rolland, @jojodecayz , and Matt Miller who will walk you through what's new and answer questions live.

model-releasescomfyui--x
26 Jun 2026
Model Releases

KARLA: Knowledge-base Augmented Retrieval for Language Models

DGX agent

arXiv:2606.26807v1 Announce Type: new Abstract: We propose a new method that allows an LLM to automatically pull in factual knowledge from a knowledge base during token generation. This means that (1)

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Know2Guess: A Contamination-Aware Multi-Zone Benchmark for Knowledge-Boundary Evaluation in Large Language Models

DGX agent

arXiv:2606.26101v1 Announce Type: cross Abstract: Reliable evaluation of large language models should separate supported answering from unsupported guessing without conflating either with data contami

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

LA4VLA: Learning to Act without Seeing via Language-Action Pretraining

DGX agent

arXiv:2606.27295v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly pretrained on robot demonstrations by jointly mapping visual observations and language instructions to

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

Latent Diffusion Posterior Sampling with Surrogate Likelihood Guidance for PDE Inverse Problems

DGX agent

arXiv:2606.26592v1 Announce Type: cross Abstract: We propose latent-space diffusion posterior sampling (L-DPS), an approximate Bayesian framework for high-dimensional inverse problems governed by part

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Layer-Specific Prompt Fusion Discovery via Differentiable Search in Vision Foundation Models

DGX agent

arXiv:2606.26379v1 Announce Type: new Abstract: Visual prompt tuning has emerged as a parameter-efficient fine-tuning approach for adapting large-scale Vision Transformers (ViTs) to downstream tasks.

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Layered Outer-Loop Control for Disturbance-Robust Multi-Waypoint UAV Arrival

DGX agent

arXiv:2606.26315v1 Announce Type: new Abstract: Disturbance-robust UAV position control is easy to demonstrate in benign simulations but much harder to make fast in approach, well behaved near the tar

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

LCAi: Life Cycle Assessment with big data fusion and retrieval-augmented generation-assisted interpretation

DGX agent

arXiv:2606.26857v1 Announce Type: new Abstract: The interpretation phase of life cycle assessment often lacks structured mechanisms for translating quantified improvement opportunities addressing envi

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Learning Language-Driven Sequence-Level Modal-Invariant Representations for Video-Based Visible-Infrared Person Re-Identification

DGX agent

arXiv:2601.12062v2 Announce Type: replace Abstract: The core of video-based visible-infrared person re-identification (VVI-ReID) lies in learning sequence-level modal-invariant representations across

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Learning Long-Range Dependencies with Temporal Predictive Coding

DGX agent

arXiv:2602.18131v2 Announce Type: replace Abstract: Temporal Predictive Coding provides a layer-local, parallelisable mechanism for learning in recurrent systems, making it an attractive candidate for

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Learning Motion Feasibility from Point Clouds in Cluttered Environments

DGX agent

arXiv:2606.26700v1 Announce Type: cross Abstract: Motion feasibility prediction plays a central role in robotics, particularly in task and motion planning and manipulation. A major bottleneck for this

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Learning to Recover Task Experts from a Multi-Task Merged Model

DGX agent

arXiv:2606.26902v1 Announce Type: new Abstract: Multi-task model merging aims to consolidate several task-specific experts into a unified model, yet static merging consistently suffers from parameter

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Learning to Select Maximum Clique Algorithms: From Traditional Machine Learning to a Dual-Channel Hybrid Neural Architecture

DGX agent

arXiv:2508.08005v4 Announce Type: replace-cross Abstract: The Maximum Clique Problem (MCP) is an NP-hard problem with wide-ranging applications in fields such as bioinformatics, network science, and s

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Letter: the US lifts its block on Mythos 5, allowing Anthropic to release it to more than 100 US institutions; sources: talks about Fable 5 are ongoing (Semafor)

DGX agent

Semafor: Letter: the US lifts its block on Mythos 5, allowing Anthropic to release it to more than 100 US institutions; sources: talks about Fable 5 are ongoing — THE SCOOP — The US government Friday

model-releasestechmeme
26 Jun 2026
Model Releases

Life After Benchmark Saturation: A Case Study of CORE-Bench

DGX agent

arXiv:2606.26158v1 Announce Type: new Abstract: When a benchmark's accuracy saturates, it is often retired and replaced with a more challenging version. We show that this approach privileges accuracy

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Limited Reference, Reliable Generation: A Two-Component Framework for Tabular Data Generation in Low-Data Regimes

DGX agent

arXiv:2509.09960v2 Announce Type: replace-cross Abstract: Synthetic tabular data generation is increasingly essential in machine learning, supporting downstream applications when real-world, high-qual

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

LiMoDE: Rethinking Lifelong Robot Manipulation from a Mixture-of-Dynamic-Experts Perspective

DGX agent

arXiv:2606.26183v1 Announce Type: cross Abstract: Building a generalist robot that can leverage prior knowledge for continuous task adaptation remains a significant challenge. Previous works alleviate

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Liquid Fusion of Heterogeneous Representations Towards General Salient Object Detection

DGX agent

arXiv:2606.26849v1 Announce Type: new Abstract: General Salient Object Detection (SOD) aims to identify and segment visually interesting objects from uni-modality or multi-modality scenes, recently ad

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing

DGX agent

arXiv:2606.26740v1 Announce Type: new Abstract: Streaming video editing has made rapid progress, yet practical deployment is still limited by two core issues: maintaining stable backgrounds and non-ed

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

LMs as Task-Specific Knowledge Bases: An Interpretability Analysis

DGX agent

arXiv:2606.27237v1 Announce Type: new Abstract: Language models (LMs) capture large amounts of factual knowledge applicable to a wide range of tasks, motivating the view of their parameters as a knowl

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds

DGX agent

arXiv:2606.26964v1 Announce Type: new Abstract: As embodied AI and world models increasingly operate in dynamic 3D environments, visual perception must move beyond passively interpreting given observa

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Low Resource Multimodal Translation of Nepali Spoken Words into Emotion-Conditioned Sign Language Avatars

DGX agent

arXiv:2606.26107v1 Announce Type: cross Abstract: Sign language communication systems, that integrate emotional expression remain underexplored, particularly for low-resource languages. This pilot stu

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Mask to Concept: Auto-Promptable SAM3 via Efficient Test-Time Concept Embedding Search for Few-Shot Annotation

DGX agent

arXiv:2606.26711v1 Announce Type: new Abstract: Transforming foundation segmentation models from human-prompted tools into auto-promptable annotators is critical for scalable medical data annotation.

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Mean-Field PhiBE: Continuous-Time Mean-Field Reinforcement Learning from Discrete-Time Data

DGX agent

arXiv:2606.26498v1 Announce Type: cross Abstract: This paper addresses model-free continuous-time mean-field control in a setting where the population dynamics evolve continuously according to an unkn

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Memory Depth, Not Memory Access: Selective Parametric Consolidation for Long-Running Language Agents

DGX agent

arXiv:2606.26806v1 Announce Type: new Abstract: Long-running language agents need more than memory access. Retrieval systems can fetch past facts at query time, but they do not decide which experience

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

MetaboNet-Bench: A Multi-modal Benchmark for Glucose Forecasting in Type 1 Diabetes

DGX agent

arXiv:2606.18640v2 Announce Type: replace Abstract: Glucose forecasting algorithms are an important aspect of glycemic control management in type 1 diabetes. So far, the research community has develop

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

MKG-RAG-Bench: Benchmarking Retrieval in Multimodal Knowledge Graph-Augmented Generation

DGX agent

arXiv:2606.26458v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) over knowledge graphs has emerged as a promising approach for grounding large language models, yet existing benchma

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

MLFFM-SegDiff: A Multi-Level Feature Fusion Diffusion Model for Skin Lesion Segmentation

DGX agent

arXiv:2606.26712v1 Announce Type: cross Abstract: Skin lesion segmentation is a key task in computer-aided dermatological diagnosis, where accuracy directly impacts downstream analysis and disease cla

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

> mythos is so good at cyber it can't be released also > mythos can't detect 20k fraudulent chinese accounts attacking it

DGX agent

This post discusses apparent contradictions in claims about Mythos' cybersecurity capabilities, suggesting tension between assertions that it excels at cyber defense versus reports that it failed to d

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

NavIsaacLab: Generating Realistic Crowd via Parallel Robot Learning for Benchmarking Human-aware Navigation

DGX agent

arXiv:2606.26265v1 Announce Type: new Abstract: Robot autonomous navigation that accounts for surrounding human activities is crucial for ensuring both safety and natural human-robot interaction in re

model-releasesarxiv-cs-ro
26 Jun 2026
← Previous
1…165166167168169…475
Next →