AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
Model Releases

RadiomicNet: A Hybrid Radiomics-Guided Lightweight Architecture for Interpretable Medical Image Segmentation

DGX agent

arXiv:2607.02185v1 Announce Type: cross Abstract: Deep learning has achieved remarkable performance in medical image segmentation, yet it suffers from critical limitations: mathematical intractability

model-releasesarxiv-cs-ai
3 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Reasoning LLM Improves Speaker Recognition in Long-form TV Dramas

DGX agent

arXiv:2607.02504v1 Announce Type: cross Abstract: Long-form TV dramas present a formidable challenge for comprehensive video understanding, where deciphering complex storyline often relies on extbf{sp

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration

DGX agent

arXiv:2603.06001v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models enable robots to perform manipulation tasks directly from natural language instructions and are increasing

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Robust for the Wrong Reasons: The Representational Geometry of LLM Robustness to Science Skepticism

DGX agent

arXiv:2607.01951v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly consulted on contested scientific questions, raising the concern that they will sycophantically retreat

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Rocket Report: Indian startup nears first launch; SpaceX's millenary milestone

DGX agent

India's first space unicorn is preparing for its first orbital launch , with Skyroot Aerospace developing the Vikram series of launch vehicles . In 2025, SpaceX posted a net loss of $4.9 billion , hig

model-releasesars-technica
3 Jul 2026
Model Releases

RusFinChain: A Russian Benchmark for Verifiable Chain-of-Thought Reasoning in Finance with Fuzzy-Aligned Evaluation

DGX agent

arXiv:2607.01388v1 Announce Type: new Abstract: Multi-step symbolic reasoning is essential for robust financial analysis, yet most benchmarks neglect intermediate reasoning steps. FINCHAIN introduced

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

SAB-LVLM: Significance-Aware Binarization for Large Vision-Language Models

DGX agent

arXiv:2607.01876v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in multimodal understanding, yet their enormous parameter scale and cross-modal

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Targeted Embedding Exploit via Refinement

DGX agent

arXiv:2607.01859v1 Announce Type: new Abstract: Safety training for large language models (LLMs) is conducted predominantly in English, leaving uncertain how well safety mechanisms generalize to low-r

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification

DGX agent

arXiv:2607.01793v1 Announce Type: new Abstract: LLM agents increasingly perform autonomous actions through external tools, leading to complex and evolving safety risks. However, existing safety testin

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Scaling Trends for Lie Detector Oversight in Preference Learning

DGX agent

arXiv:2607.01567v1 Announce Type: new Abstract: Deceptive behavior in LLMs is costly to monitor and prevent, motivating approaches such as Scalable Oversight via Lie Detectors (SOLiD) (Cundy & Gleave,

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling

DGX agent

arXiv:2607.01612v1 Announce Type: new Abstract: Training large language models (LLMs) with reinforcement learning (RL) has significantly advanced their performance on reasoning and question-answering

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SCAPE: Accurate and Efficient LLM Training with Extreme Sparse Communication

DGX agent

arXiv:2607.01678v1 Announce Type: new Abstract: Communication increasingly dominates the cost of Large Language Model (LLM) pre-training, especially under data-parallel and sharded training schemes, w

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Self-Gating Attention for Efficient Time Series Forecasting

DGX agent

arXiv:2607.02344v1 Announce Type: cross Abstract: Transformer architectures have shown strong potential in time series forecasting, where multi-head self-attention is widely used to capture temporal d

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Separating Expert Retention from Autonomous Source Inference in Raw-ECG-Replay-Free Continual ECG Deployment

DGX agent

arXiv:2607.01674v1 Announce Type: new Abstract: In multi-source ECG deployment, models may need to incorporate new data sources when earlier raw ECGs cannot be retained or replayed. Freezing a pretrai

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SimWorlds: A Multi-Agent System for Dynamic 3D Scene Creation

DGX agent

arXiv:2607.01766v1 Announce Type: new Abstract: LLM agents are increasingly used to translate natural language into 3D scenes in a procedural way, but existing systems focus on static output. Dynamic

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Some notes from @aiDotEngineer world fair: > the energy was incredible. it's magical to have a large group of smart, hungry, technical, and …

DGX agent

Some notes from @aiDotEngineer world fair: > the energy was incredible. it's magical to have a large group of smart, hungry, technical, and driven people learning from each other under one roof. > the

model-releasesswyx--x
3 Jul 2026
Model Releases

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was th…

DGX agent

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was the year of realizing that autonomy without structure creates

model-releasesyohei-nakajima--x
3 Jul 2026
Model Releases

Sources: Alibaba has banned employees from using Claude Code and asked them to remove all Claude models from their work computers, citing security concerns (The Information)

DGX agent

The Information: Sources: Alibaba has banned employees from using Claude Code and asked them to remove all Claude models from their work computers, citing security concerns — Alibaba Group has banned

model-releasestechmeme
3 Jul 2026
Model Releases

Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters

DGX agent

arXiv:2607.01893v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by drafting a block of tokens that the target model verifies left-to-right, committing only t

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Spectral Imbalance Causes Forgetting in Low-Rank Continual Adaptation

DGX agent

arXiv:2602.00722v2 Announce Type: replace Abstract: Parameter-efficient continual learning aims to adapt pre-trained models to sequential tasks without forgetting previously acquired knowledge. Most e

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

SPLIT: Cross-Lingual Empathy and Cultural Grounding in English and Ukrainian LLM Responses

DGX agent

arXiv:2607.02049v1 Announce Type: cross Abstract: Large Language Models are increasingly deployed in emotional-support contexts and crisis-related situations. Nevertheless, their cross-lingual abiliti

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SPOT: Spatio-Temporal Obstacle-free Trajectory Planning for UAVs in Unknown Dynamic Environments

DGX agent

arXiv:2602.01189v3 Announce Type: replace Abstract: We address the problem of reactive motion planning for quadrotors operating in unknown environments with dynamic obstacles. Our approach leverages a

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics

DGX agent

arXiv:2510.09517v2 Announce Type: replace Abstract: Despite rapid advances in large language models (LLMs), statistical reasoning remains underrepresented in existing LLM benchmarks, which often do no

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Steerability via constraints: a substrate for scalable oversight of coding agents

DGX agent

arXiv:2607.02389v1 Announce Type: new Abstract: Coding agents are capable; human oversight is the bottleneck. Unconstrained agents introduce security risks, erode codebase scalability, and make human

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Structured Gaussian Processes for Uncertainty-Aware Classification of High-Dimensional, Small-Sampled Omics Data

DGX agent

arXiv:2607.02103v1 Announce Type: cross Abstract: Classifying heterogeneous omics data remains a fundamental challenge in computational biology, particularly in high-dimensional, small-sample settings

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

TestEvo-Bench: An Executable and Live Benchmark for Test and Code Co-Evolution

DGX agent

arXiv:2607.02469v1 Announce Type: cross Abstract: Software tests and code evolve together: a code change should be followed by new or updated tests that record the new software behavior. Yet existing

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Text-Driven 3D Indoor Scene Synthesis in Non-Manhattan Environments

DGX agent

arXiv:2607.02407v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in 3D indoor synthesis for Manhattan environments. However, existing methods ofte

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

The Eticas AI Risk Taxonomy: Open Infrastructure for Operationalizing AI Audits

DGX agent

arXiv:2607.02201v1 Announce Type: cross Abstract: The rapid deployment of AI systems across high-stakes domains has created urgent demand for standardized evaluation, yet the field remains fragmented

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

the scroll feature in the claude code CLI is really nice

DGX agent

Jerry Liu highlights the scroll feature in the Claude Code CLI as a beneficial functionality. The post suggests this feature improves the user experience when working with the Claude Code command-line

model-releasesjerry-liu--x
3 Jul 2026
Model Releases

The team at @vercel recently released the Eve agent framework, so we built a template that integrates LiteParse with it🦙 The template provi…

DGX agent

The team at @vercel recently released the Eve agent framework, so we built a template that integrates LiteParse with it🦙 The template provides a set of read-only filesystem tools that let Eve resolve

model-releasesjerry-liu--x
3 Jul 2026
Model Releases

The Wiola Architecture for Efficient Small Language Models

DGX agent

arXiv:2607.01394v1 Announce Type: new Abstract: We present Wiola, a fully original Small Language Model (SLM) architecture built from first principles, sharing no structural lineage with any existing

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

This is actually useful. LangChain just released OpenWiki. It's an open-source agent that creates a wiki for your codebase, connects it to y…

DGX agent

This is actually useful. LangChain just released OpenWiki. It's an open-source agent that creates a wiki for your codebase, connects it to your coding agent, and keeps it updated as your repo changes.

model-releasesharrison-chase--x
3 Jul 2026
Model Releases

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as…

DGX agent

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as a Google replacement, for homework “help,” etc. It is someo

model-releasesethan-mollick--x
3 Jul 2026
Model Releases

Token Geometry

DGX agent

arXiv:2607.01455v1 Announce Type: cross Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them.

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Towards a Phonology-Informed Evaluation of Multilingual TTS

DGX agent

arXiv:2607.01965v1 Announce Type: new Abstract: Neural TTS systems can sound natural across languages, but naturalness does not guarantee the preservation of sound contrasts that distinguish words fro

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Towards Load-Aware Prefill Deflection for Disaggregated LLM Serving

DGX agent

arXiv:2607.02043v1 Announce Type: cross Abstract: Disaggregated LLM serving runs prefill and decode on separate GPU pools to keep the two phases from interfering. In practice, this creates a new asymm

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Towards Robustness against Typographic Attack with Training-free Concept Localization

DGX agent

arXiv:2607.02494v1 Announce Type: cross Abstract: Models trained via Contrastive Language-Image Pretraining (CLIP) serve as the foundational vision encoders for most modern Large Vision Language Model

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

TUDUM: A Turkish-Thinking Reasoning Pipeline for Qwen3.5-27B

DGX agent

arXiv:2607.01927v1 Announce Type: cross Abstract: This paper presents TUDUM (Turkce Dusunen Uretken Model), a project pipeline for adapting a Qwen-family 27B thinking model toward Turkish reasoning. T

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

TurnNat: Automatic Evaluation of Turn-Taking Naturalness in Dyadic Spoken Dialogue

DGX agent

arXiv:2607.01345v1 Announce Type: cross Abstract: Turn-taking naturalness is central to full-duplex spoken dialogue systems, yet its automatic evaluation remains limited. Existing evaluations often re

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

DGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Understanding Agent-Based Patching of Compiler Missed Optimizations

DGX agent

arXiv:2607.02370v1 Announce Type: cross Abstract: Compiler missed optimizations refer to cases in which compilers failed to optimize certain code. It takes many compiler developers' efforts to impleme

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few…

DGX agent

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few months is *Qwen 27b*. Our ML/AI engineering teams are have

model-releasesclem-delangue--x
3 Jul 2026
Model Releases

VisionAId: An Offline-First Multimodal Android Assistant for People with Visual Impairment, Featuring Personalized Object Retrieval

DGX agent

arXiv:2607.02371v1 Announce Type: cross Abstract: Over 285 million people worldwide live with a visual impairment, for whom everyday tasks such as avoiding obstacles, locating personal belongings, rec

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer

DGX agent

arXiv:2512.11891v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities in generalizing across diverse robotic manipulation tasks. However, de

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

WARP: Weight-Space Analysis for Recovering Training Data Portfolios

DGX agent

arXiv:2607.01686v1 Announce Type: new Abstract: Foundation models are routinely released to the public, yet the data recipes used to train them -- such as domain mixture weights that determine how dif

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

We analyzed GLM 5.2 vs Sonnet 5 for software engineering tasks using DeepSWE. GLM 5.2 gets you ~80% of Sonnet 5's capability at ~20% of the …

DGX agent

We analyzed GLM 5.2 vs Sonnet 5 for software engineering tasks using DeepSWE. GLM 5.2 gets you ~80% of Sonnet 5's capability at ~20% of the price. More insights in the thread! Deepdive: Sonnet 5 and G

model-releasestogether-ai--x
3 Jul 2026
Model Releases

we distilled 2.3M Claude Fable 5 reasoning traces into Qwen3-4B - 100% self-consistency @ 512 samples - 0.00 bits output entropy - zero hall…

DGX agent

we distilled 2.3M Claude Fable 5 reasoning traces into Qwen3-4B - 100% self-consistency @ 512 samples - 0.00 bits output entropy - zero hallucination variance turns out the student is not bounded by t

model-releasesclem-delangue--x
3 Jul 2026
Model Releases

Yes they can move and dance and stuff You can assume they will talk and sing and more as good as anyone https://x.com/ubtechrobotics/status/…

DGX agent

Yes they can move and dance and stuff You can assume they will talk and sing and more as good as anyone https://x.com/ubtechrobotics/status/2072651508710285419?s=46 UBTECH Launches UWORLD U1 — The Wor

model-releasesemad-mostaque--x
3 Jul 2026
← Previous
1…134135136137138…471
Next →