AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,585 results
Model Releases

Great paper on managing agent skills. Skill libraries keep growing, and picking the right skills has become a bottleneck for coding agents. …

DGX agent

Great paper on managing agent skills. Skill libraries keep growing, and picking the right skills has become a bottleneck for coding agents. The defaults are to expose the agent to the whole skill coll

model-releasesdair-ai--x
1 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

guess who’s back, back again

DGX agent

guess who’s back, back again We’ve received notice that the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5. We'll begin restoring access tomorrow, and will share an u

model-releasesyohei-nakajima--x
1 Jul 2026
Model Releases

Happy Canada Day!! 🇨🇦 We're so proud to be founded and growing in Canada. Let's keep building💡

DGX agent

Cohere, an AI company, celebrates Canada Day and expresses pride in being founded and operating in Canada, while affirming their commitment to continued growth and development in the country. The post

model-releasescohere--x
1 Jul 2026
Model Releases

Have seen some questions about the updated classifiers and wanted to clarify. As with the original classifiers, a small fraction of routine …

DGX agent

Have seen some questions about the updated classifiers and wanted to clarify. As with the original classifiers, a small fraction of routine coding and debugging tasks will be flagged and fall back to

model-releasesthariq--x
1 Jul 2026
Model Releases

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents

DGX agent

arXiv:2606.31179v1 Announce Type: new Abstract: As AI agents become increasingly capable of complex, long-horizon reasoning, rigorous and holistic evaluation is essential for measuring progress toward

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

High-Speed Vision-Based Flight in Clutter with Safety-Shielded Reinforcement Learning

DGX agent

arXiv:2602.08653v2 Announce Type: replace Abstract: Quadrotor unmanned aerial vehicles (UAVs) are increasingly deployed in complex missions that demand reliable autonomous navigation and robust obstac

model-releasesarxiv-cs-ro
1 Jul 2026
Model Releases

HSDF-Lane: Height-Aligned Signed Distance Field with Semantic Lane Prior for 3D Lane Detection

DGX agent

arXiv:2606.31172v1 Announce Type: new Abstract: Monocular 3D lane detection plays a critical role in autonomous driving, yet recovering reliable 3D geometry from a single image remains challenging due

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Hugging Face and Cerebras bring Gemma 4 to real-time voice AI

DGX agent

Hugging Face and Cerebras have collaborated to integrate Google's Gemma 4 model with real-time voice AI capabilities, enabling faster speech processing and voice interactions. This integration likely

model-releaseshugging-face
1 Jul 2026
Model Releases

Human-Agent Collaborative Paper-to-Page Crafting

DGX agent

arXiv:2510.19600v2 Announce Type: replace-cross Abstract: In the quest for scientific progress, communicating research is as vital as the discovery itself. Yet, researchers are often sidetracked by th

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

HVPNet: A Bio-Inspired Network for General Salient and Camouflaged Object Detection

DGX agent

arXiv:2606.31496v1 Announce Type: new Abstract: In recent years, most research on multimodal salient object detection (SOD) and camouflaged object detection (COD) typically aims to improve performance

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

HyPOLE: Hyperproperty-Guided Multi-Agent Reinforcement Learning under Partial Observation

DGX agent

arXiv:2606.30966v1 Announce Type: new Abstract: Formal specification is a powerful tool to guide the learning process and provides significant advantages over reward shaping: (1) mathematical rigor; (

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Improving LLM Reasoning with Homophily-aware Structural and Semantic Text-Attributed Graph Compression

DGX agent

arXiv:2601.08187v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated promising capabilities in Text-Attributed Graph (TAG) understanding. Recent studies typically focus o

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Indi-RomCoM: Code-Mixed Benchmark for Evaluating LLMs on Romanized Indic-English Instructions

DGX agent

arXiv:2606.30790v1 Announce Type: cross Abstract: Romanized Code Mixing (RCM), where bilingual speakers fluidly blend local languages with English in Roman script, has emerged as the dominant form of

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Introducing ZCode, the official development environment for GLM-5.2 - GLM Coding Plan subscribers: now 1.5x usage quota in ZCode - BYOK supp…

DGX agent

Introducing ZCode, the official development environment for GLM-5.2 - GLM Coding Plan subscribers: now 1.5x usage quota in ZCode - BYOK supported: works with your existing subscriptions and APIs - Ava

model-releaseszhipu-ai--x
1 Jul 2026
Model Releases

JacobianAvatar: Temporally Consistent Semi-rigid Avatar Reconstruction from a Monocular Video

DGX agent

arXiv:2606.31115v1 Announce Type: new Abstract: Generating realistic human avatars in complex motions--such as clothing dynamics--requires modeling of global and local deformations which remains chall

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Jamf launches Beacon threat hunting service for enterprise Mac environments

DGX agent

Apple enterprise management firm Jamf Holding Corp. today launched Beacon by Jamf Threat Labs, a premium threat hunting service that puts the company’s research and detection engineers directly inside

model-releasessiliconangle
1 Jul 2026
Model Releases

JL1-CC&QA: Extending the JL1-CD Benchmark with Change Captioning and Question Answering

DGX agent

arXiv:2606.31745v1 Announce Type: cross Abstract: Remote sensing change detection (CD) traditionally focuses on pixel-level binary segmentation, which identifies where changes occur but neither what n

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

JUST IN: Portugal launches its first open-source AI model, joining Europe’s push for “tech sovereignty”

DGX agent

Portugal has launched its first open-source AI model as part of Europe's broader initiative to achieve 'tech sovereignty' and reduce dependence on non-European AI systems. This development aligns with

model-releasesclem-delangue--x
1 Jul 2026
Model Releases

Knowledge Distillation from Large Reasoning Models to Compact Student Models: A Case Study on the John O Bryan Mathematics Competition

DGX agent

arXiv:2606.31048v1 Announce Type: cross Abstract: This paper investigates knowledge distillation from a large reasoning model (DeepSeek-R1) to a compact student model (Qwen2.5-7B). Using historical pr

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Labimus: A Simulation and Benchmark for Humanoid Dexterous Manipulation in Chemical Laboratory

DGX agent

arXiv:2606.31037v1 Announce Type: new Abstract: Laboratory automation has made remarkable progress through robotic platforms and AI-driven scientific reasoning. However, many laboratory operations (e.

model-releasesarxiv-cs-ro
1 Jul 2026
Model Releases

Learning from Failure: Inference-Time Self-Improvement for Computer-Use Agents

DGX agent

arXiv:2606.31270v1 Announce Type: cross Abstract: Computer-use agents, which leverage multimodal large language models (MLLMs) to operate computers and complete tasks, have attracted significant atten

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Learning to Deny: Action Denial in Multimodal Large Language Models

DGX agent

arXiv:2606.31187v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have rapidly advanced video understanding, achieving strong zero-shot and few-shot recognition across standard

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

LightOnOCR: A 1B End-to-End Multilingual Vision-Language Model for State-of-the-Art OCR

DGX agent

arXiv:2601.14251v2 Announce Type: replace Abstract: We present LightOnOCR-2-1B, a 1B-parameter end-to-end multilingual vision--language model that converts document images (e.g., PDFs) into clean, nat

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

LLM-as-a-judge validity in physics assessment depends more on the task than the model

DGX agent

arXiv:2603.14732v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly considered for automated assessment and feedback, understanding when LLM marking is valid is

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

LLMs are stuck in a groupthink groove. This startup is trying to get them out.

DGX agent

Let’s start with a game. Open up your chatbot of choice—Claude, ChatGPT, Gemini—and type “Give me a random number between 1 and 10.” You’re going to get 7. Almost always. Now type “Another” and you’ll

model-releasesmit-tech-review
1 Jul 2026
Model Releases

LOPA: Enhancing Spoken Language Assessment via Latent Ordinal Prototype Alignment

DGX agent

arXiv:2606.31310v1 Announce Type: new Abstract: Fueled by increasing model scale and multimodal inputs, Multimodal Large Language Models (MLLMs) have emerged as a promising paradigm for Spoken Languag

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Loved the chat between @trq212 @_catwu @simonw at AI Eng summit. My top 13 takeaways from their session -> 1. Engineers should become better…

DGX agent

Loved the chat between @trq212 @_catwu @simonw at AI Eng summit. My top 13 takeaways from their session -> 1. Engineers should become better at product/business sense. 2. Don't worry about major rewri

model-releasesswyx--x
1 Jul 2026
Model Releases

LuxEmo: Expressive Text-to-Speech Corpus for Luxembourgish

DGX agent

arXiv:2606.31947v1 Announce Type: new Abstract: State-of-the-art speech datasets predominantly focus on widely spoken languages, often overlooking low-resource languages such as Luxembourgish, which r

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

M3CoTBench: Benchmark Chain-of-Thought of MLLMs in Medical Image Understanding

DGX agent

arXiv:2601.08758v4 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) reasoning has proven effective in enhancing large language models by encouraging step-by-step intermediate reasoning, a

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Measuring & Mitigating Over-Alignment for LLMs in Multilingual Criminal Law Courts

DGX agent

arXiv:2606.23375v2 Announce Type: replace-cross Abstract: While the wider applicability of LLMs in the legal field is currently debated due to their reliability and the gravity of any errors, narrow u

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments

DGX agent

arXiv:2606.31966v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have strong potential as embodied agents, but their ability to collaborate in visually grounded enviro

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Medical Image Spatial Grounding with Semantic Sampling

DGX agent

arXiv:2603.14579v3 Announce Type: replace Abstract: Vision language models (VLMs) have shown significant promise in visual grounding for images as well as videos. In medical imaging research, VLMs rep

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Mind the Residual Gap: Probabilistic Downscaling under Real-World Bias

DGX agent

arXiv:2606.30821v1 Announce Type: new Abstract: Probabilistic downscaling is the task of modeling the conditional distribution of high-resolution fields given coarse inputs, and is a central challenge

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

MIRTH: Mutual-Information Reasoning with Temporal Hubs for Vision-Language-Action Agents

DGX agent

arXiv:2606.31167v1 Announce Type: cross Abstract: VLA models have emerged as a powerful paradigm for transferring semantic knowledge from web-scale data to physical robotic control. However, current s

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Mixture-of-Control: State-Aware Fine-Tuning for Transformer-based Models

DGX agent

arXiv:2606.31397v1 Announce Type: cross Abstract: State-based fine-tuning has emerged as a compelling alternative to weight-based adaptation for transformers, updating lightweight controls into states

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Modality-Driven Search with Holistic Trace Judging for ARC-AGI-2

DGX agent

arXiv:2606.31543v1 Announce Type: new Abstract: Large language models can produce fluent, internally coherent reasoning traces for abstract reasoning tasks while still being confidently wrong - making

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Modeling Cell-Cycle-Aware Single-Cell Drug Perturbation Responses

DGX agent

arXiv:2606.30695v1 Announce Type: cross Abstract: Single-cell drug perturbation models should predict not only transcriptional response magnitude, but also whether a treatment alters the proliferative

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Moral Safety in LLMs: Exposing Performative Compliance with Puzzled Cues

DGX agent

arXiv:2606.31644v1 Announce Type: new Abstract: As large language models take on morally consequential roles in healthcare, legal, and hiring contexts, we need to examine whether their ethical behavio

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Multi-GPU kernels are the real test for coding models. Today at @aiDotEngineer, @simran_s_arora shared ParallelKernelBench, an open-source b…

DGX agent

Multi-GPU kernels are the real test for coding models. Today at @aiDotEngineer, @simran_s_arora shared ParallelKernelBench, an open-source benchmark for evaluating whether LLMs can write fast CUDA ker

model-releasestogether-ai--x
1 Jul 2026
Model Releases

Multimodal Benchmark for Safety Assessment in Industrial Inspection Scenarios

DGX agent

arXiv:2601.21173v2 Announce Type: replace-cross Abstract: With the rapid development of industrial intelligence and unmanned inspection, reliable perception and safety assessment for AI systems in com

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining

DGX agent

arXiv:2509.10406v4 Announce Type: replace Abstract: Pretraining transformers on long sequences (entire code repositories, collections of related documents) is bottlenecked by quadratic attention costs

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

MultiUAV-Plat: An LLM-Oriented Platform, Benchmark and Framework for Multi-UAV Collaborative Task Planning

DGX agent

arXiv:2606.31073v1 Announce Type: new Abstract: Large language models (LLMs) provide a promising interface for high-level robotic task planning, but their use in multi-UAV collaboration remains diffic

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

My prediction: the excitement for Fable 5 will wear off really fast. Reposting this to help those who will be extremely disappointed after t…

DGX agent

My prediction: the excitement for Fable 5 will wear off really fast. Reposting this to help those who will be extremely disappointed after they play with Fable 5 and run out of tokens or can't do much

model-releasesdair-ai--x
1 Jul 2026
Model Releases

Nemotron 3 Ultra is taking off on Together AI 🚀 In just days, it's climbed to 35B tokens/day on @OpenRouter. That's the community voting wi…

DGX agent

Nemotron 3 Ultra is taking off on Together AI 🚀 In just days, it's climbed to 35B tokens/day on @OpenRouter. That's the community voting with their tokens for open models that are fast, efficient, and

model-releasestogether-ai--x
1 Jul 2026
Model Releases

Neurologyca launches research labs to help frontier AI models understand people

DGX agent

Neurologyca, a company working on human-centric artificial intelligence, announced the launch of a research division on Tuesday dedicated to advancing AI models’ ability to understand and adapt to peo

model-releasessiliconangle
1 Jul 2026
Model Releases

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs

DGX agent

arXiv:2606.31933v1 Announce Type: new Abstract: We introduce VidPair-Halluc, a new benchmark for evaluating video hallucination in large video models (LVMs) under rigorous and controlled conditions. U

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Nonlinearity-Aware LoRA: Structured Gate Adaptation under Low-Rank Constraints

DGX agent

arXiv:2606.31717v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) is commonly viewed as an update-space approximation to full fine-tuning, yet this view is incomplete for self-gated Transform

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

NURBS Splatting: A Unified Differentiable Rendering Framework for Vector Graphics

DGX agent

arXiv:2606.31764v1 Announce Type: cross Abstract: Differentiable rendering of planar rational splines remains largely underexplored, despite their widespread use in vector graphics and design. Existin

model-releasesarxiv-cs-cv
1 Jul 2026
← Previous
1…142143144145146…471
Next →