AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,637 results
Model Releases

BioBlue: Systematic runaway-optimiser-like LLM failure modes on biologically and economically aligned AI safety benchmarks for LLMs with simplified observation format

DGX agent

arXiv:2509.02655v3 Announce Type: replace-cross Abstract: Many AI alignment discussions of 'runaway optimisation' focus on RL agents: unbounded utility maximisers that over-optimise a proxy objective

model-releasesarxiv-cs-ai
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Biodefense in the Intelligence Age

DGX agent

This document likely examines how artificial intelligence and advanced intelligence capabilities can be applied to biodefense strategies, including disease surveillance, threat detection, and pandemic

model-releasesopenai
4 Jun 2026
Research

Boosting Self-Consistency with Ranking

DGX agent

arXiv:2606.05054v1 Announce Type: new Abstract: Self-consistency improves large language models by sampling multiple reasoning paths and selecting the most frequent answer, but majority voting often f

researcharxiv-cs-cl
4 Jun 2026
Agents

CADENCE: Predicting Realized MAPF Execution Time Beyond Sum of Costs

DGX agent

arXiv:2606.04746v1 Announce Type: new Abstract: Multi-Agent Path Finding (MAPF) algorithms are increasingly used to plan motion for robot teams in industrial warehouses and robotic shared workspaces,

agentsarxiv-cs-ro
4 Jun 2026
Tutorials

Can Crowdsourcing Survive the LLM Era? A Community Survey on Human Data Collection

DGX agent

arXiv:2606.04924v1 Announce Type: new Abstract: The widespread use of Large Language Models (LLMs) as writing tools challenges the validity of crowdsourced data, as crowdworkers may outsource tasks to

tutorialsarxiv-cs-cl
4 Jun 2026
Safety

Certified Neural Approximations of Nonlinear Dynamics

DGX agent

arXiv:2505.15497v3 Announce Type: replace Abstract: Neural networks hold great potential to act as approximate models of nonlinear dynamical systems, with the resulting neural approximations enabling

safetyarxiv-cs-lg
4 Jun 2026
Model Releases

Continual Visual and Verbal Learning Through a Child's Egocentric Input

DGX agent

arXiv:2606.05115v1 Announce Type: cross Abstract: Children learn the meanings of words from a continuous, temporally structured stream of egocentric experience. Recent work shows that neural networks

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

CoPark: Learning Reactive Parking via Self-Play

DGX agent

arXiv:2606.04149v1 Announce Type: new Abstract: Learning a single policy that reaches a goal with high geometric precision while interacting safely with nearby agents poses conflicting objectives. Pre

model-releasesarxiv-cs-ro
4 Jun 2026
Safety

Crafting Your Evolving Dreams: Concept-Incremental Versatile Customization

DGX agent

arXiv:2606.04797v1 Announce Type: new Abstract: Custom diffusion models (CDMs) have garnered significant interest owing to their remarkable capacity for generating personalized concepts. However, the

safetyarxiv-cs-cv
4 Jun 2026
Model Releases

Curvature-aware dynamic precision approach for physics-informed neural networks

DGX agent

arXiv:2606.04736v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have become a promising framework for simulating partial differential equations (PDEs) by embedding physical

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities

DGX agent

arXiv:2606.04460v1 Announce Type: cross Abstract: AI has the potential to transform cybersecurity by enabling systems that can autonomously detect, analyze, and remediate software vulnerabilities. How

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

DeepSeek is becoming more popular among US enterprises as companies look for cheaper alternatives to Anthropic and OpenAI “DeepSeek takes to…

DGX agent

DeepSeek is becoming more popular among US enterprises as companies look for cheaper alternatives to Anthropic and OpenAI “DeepSeek takes top spot on 'trending' list as companies look for alternatives

model-releasesclem-delangue--x
4 Jun 2026
Research

Disentangling Answer Engine Optimization from Platform Growth: A Log-Based Natural Experiment on ChatGPT Referral Traffic

DGX agent

arXiv:2606.04362v1 Announce Type: cross Abstract: Large language model (LLM) 'answer engines' such as ChatGPT now send measurable referral traffic to the open web, and a practice analogous to search e

researcharxiv-cs-cl
4 Jun 2026
Safety

DuDi: Dual-Signal Distillation with Cross-Lingual Verbalizer

DGX agent

arXiv:2606.04694v1 Announce Type: new Abstract: Small language models (SLMs) are efficient and scalable, but their multilingual capabilities degrade severely at sub-billion scales, especially for Sout

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

Ekka: Automated Diagnosis of Silent Errors in LLM Inference

DGX agent

arXiv:2606.04594v1 Announce Type: cross Abstract: LLM serving frameworks are quickly evolving with a complex software stack and a vast number of optimizations. The rapid development process can introd

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Entity Binding Failures in Speech LLM Reasoning: Diagnosis and Chain-of-Thought Intervention

DGX agent

arXiv:2606.04474v1 Announce Type: new Abstract: Speech Large Language Models (SLLMs) underperform their text counterparts on complex reasoning. We reveal that this modality gap is not a uniform cognit

researcharxiv-cs-cl
4 Jun 2026
Model Releases

Exploring the Topology and Memory of Consensus: How LLM Agents Agree, Fragment, or Settle When Forming Conventions

DGX agent

arXiv:2606.04197v1 Announce Type: cross Abstract: How much should an LLM agent remember, and how should multi-agent systems be connected when trying to reach consensus? We show these two design choice

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up…

DGX agent

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up to 100hrs, and is confident enough to put a financial guarant

model-releasesswyx--x
4 Jun 2026
Model Releases

foom!

DGX agent

foom! Our internal data shows Claude is accelerating AI development—a possible path to recursive self-improvement, or AI autonomously building a more capable successor. It’s happening faster than we t

model-releasesemad-mostaque--x
4 Jun 2026
Safety

Formal Semantics for Agentic Tool Protocols: A Process Calculus Approach

DGX agent

arXiv:2603.24747v2 Announce Type: replace Abstract: The emergence of large language model agents capable of invoking external tools has created urgent need for formal verification of agent protocols.

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Founders Fund launches a TV-style game show featuring A-list founders and investors, including Sam Altman and Palmer Luckey, playing a game of Mafia (Tom Dotan/Newcomer)

DGX agent

Tom Dotan / Newcomer: Founders Fund launches a TV-style game show featuring A-list founders and investors, including Sam Altman and Palmer Luckey, playing a game of Mafia — Do people want to watch the

model-releasestechmeme
4 Jun 2026
Safety

From Agent Traces to Trust: Evidence Tracing and Execution Provenance in LLM Agents

DGX agent

arXiv:2606.04990v1 Announce Type: cross Abstract: Large language model (LLM)-based agents increasingly solve complex tasks by interacting with external tools, retrieval systems, memory modules, enviro

safetyarxiv-cs-ai
4 Jun 2026
Applications

From Motion Signals to Insights: A Unified Framework for Student Behavior Analysis and Feedback in Physical Education Classes

DGX agent

arXiv:2503.06525v2 Announce Type: replace-cross Abstract: Analyzing student behavior in educational scenarios is crucial for enhancing teaching quality and student engagement. Existing AI-based models

applicationsarxiv-cs-ai
4 Jun 2026
Research

From Ticks to Flows: Dynamics of Neural Reinforcement Learning in Continuous Environments

DGX agent

arXiv:2606.04275v1 Announce Type: cross Abstract: We present a novel theoretical framework for deep reinforcement learning (RL) in continuous environments by modeling the problem as a continuous-time

researcharxiv-cs-ai
4 Jun 2026
Safety

Good Reasoning Makes Good Demonstrations: Implicit Reasoning Quality Supervision via In-Context Reinforcement Learning

DGX agent

arXiv:2603.09803v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) improves reasoning in large language models but treats all correct solutions equally, potentia

safetyarxiv-cs-lg
4 Jun 2026
Industry

Grok Imagine 1.5 at rank 1

DGX agent

Grok Imagine 1.5, xAI's image generation model, has achieved the top ranking in image generation capabilities. This announcement was made by Elon Musk on X, indicating significant performance improvem

industryelon-musk--x
4 Jun 2026
Tutorials

Hadamard thought in image space

DGX agent

Hadamard thought in image space World Labs CEO Dr. Fei-Fei Li: 'The world is not made of words.' 'Language models have given machines an extraordinary command of concepts, vocabulary, and reasoning, b

tutorialselon-musk--x
4 Jun 2026
Model Releases

HD-DinoMoE: A Class-Aware Hierarchical Dual Mixture-of-Experts Network for Scleral Anomaly Segmentation in Complex Acquisition Scenarios

DGX agent

arXiv:2606.04888v1 Announce Type: new Abstract: Traditional Chinese Medicine (TCM) ocular inspection provides empirical cues for assessing scleral surface anomalies, but its clinical use remains subje

model-releasesarxiv-cs-cv
4 Jun 2026
Research

Hierarchical Space Partition for Surface Reconstruction

DGX agent

arXiv:2606.04891v1 Announce Type: new Abstract: Generating compact polygonal models from point clouds is a key problem in 3D vision and computer graphics. However, due to inherent limitations of LiDAR

researcharxiv-cs-cv
4 Jun 2026
Model Releases

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engine…

DGX agent

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engineers from NVIDIA collaborated to improve the multi-GPU perfor

model-releasesgeorgi-gerganov--x
4 Jun 2026
Model Releases

HighTide: An Agent-Curated Open-Source VLSI Benchmark Suite

DGX agent

arXiv:2606.04126v1 Announce Type: cross Abstract: We introduce HighTide, an evolving AI-assisted benchmark suite. Specifically, the contributions are: (i) a diverse open-source suite spanning multiple

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

HORIZON: Recoverability-Governed Curriculum for Physical-Domain Scaling

DGX agent

arXiv:2606.05143v1 Announce Type: new Abstract: Scaling robust robot policies requires more than broader randomization, because physical-domain experience must remain organized and learnable throughou

model-releasesarxiv-cs-ro
4 Jun 2026
Research

HYolo: An Intelligent IoT-Based Object Detection System Using Hypergraph Learning

DGX agent

arXiv:2606.04345v1 Announce Type: cross Abstract: This paper presents HYolo, an intelligent IoT-based object detection framework that integrates hypergraph learning into the YOLO architecture. Traditi

researcharxiv-cs-ai
4 Jun 2026
Model Releases

I am hooked on Dynamic Workflows! The idea of generating harnesses on the fly is so compelling that I reverse-engineered it for my agent orc…

DGX agent

I am hooked on Dynamic Workflows! The idea of generating harnesses on the fly is so compelling that I reverse-engineered it for my agent orchestrator. And then I built a monitoring dashboard (as an HT

model-releasesdair-ai--x
4 Jun 2026
Model Releases

In policy paper, OpenAI diverges from White House on AI safety

DGX agent

OpenAI Group PBC’s newly released proposal for how advanced artificial intelligence should be regulated differs slightly from the Trump administration’s executive order, also released this week. Relea

model-releasessiliconangle
4 Jun 2026
Tools

Introducing PDF to Lesson! Create interactive personalized courses from any PDF. 100% free & open source! Powered by GPT OSS on @togethercom…

DGX agent

Together AI announced a free, open-source tool called 'PDF to Lesson' that converts PDF documents into interactive, personalized courses using open-source GPT models running on Together's platform. Th

toolstogether-ai--x
4 Jun 2026
Applications

KITE: Kernelized and Information Theoretic Exemplars for In-Context Learning

DGX agent

arXiv:2509.15676v2 Announce Type: replace-cross Abstract: In-context learning (ICL) has emerged as a powerful paradigm for adapting large language models (LLMs) to new and data-scarce tasks using only

applicationsarxiv-cs-ai
4 Jun 2026
Model Releases

LCSHBench: A Multilingual, Consensus-Grounded Benchmark for Library of Congress Subject Heading Assignment

DGX agent

arXiv:2606.04382v1 Announce Type: cross Abstract: Automated subject cataloging assigns controlledvocabulary headings to bibliographic records, but LCSH has no standard public benchmark. We introduce L

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Leaving aside the question of consciousness, the Ted Chiang piece has a reasonable point about moral atrophy if you let AI make choices. But…

DGX agent

Leaving aside the question of consciousness, the Ted Chiang piece has a reasonable point about moral atrophy if you let AI make choices. But it is also interesting in light of the fact that repeated r

model-releasesethan-mollick--x
4 Jun 2026
Model Releases

Listen to the OpenAI Podcast on— Spotify https://open.spotify.com/episode/3ca5s3o53D5xcEKmKgLLGj?si=4a9a555641fa4293 Apple https://podcasts.…

DGX agent

Listen to the OpenAI Podcast on— Spotify https://open.spotify.com/episode/3ca5s3o53D5xcEKmKgLLGj?si=4a9a555641fa4293 Apple https://podcasts.apple.com/us/podcast/how-a-reasoning-model-cracked-an-80-yea

model-releasesopenai--x
4 Jun 2026
Model Releases

Look closely. There’s more in the Showcase.

DGX agent

OpenAI's developer account posted this message on X (formerly Twitter), likely encouraging developers to explore additional features, updates, or resources available in OpenAI's Showcase platform or d

model-releasesopenai--x
4 Jun 2026
Research

Low-Rank Decay for Grokking in Scale-Invariant Transformers: A Spectral-Geometric View

DGX agent

arXiv:2606.04405v1 Announce Type: cross Abstract: Modern Transformer architectures frequently employ normalization mechanisms such as RMSNorm and Query-Key Normalization, making parts of the model app

researcharxiv-cs-ai
4 Jun 2026
Model Releases

MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning

DGX agent

arXiv:2603.18577v2 Announce Type: replace Abstract: Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinica

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MemoryDocDataSet: A Benchmark for Joint Conversational Memory and Long Document Reasoning

DGX agent

arXiv:2606.04442v1 Announce Type: cross Abstract: AI systems increasingly need to combine two demanding capabilities: navigating multi-session conversation history and performing deep reading comprehe

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

MENTOR: A Metacognition-Driven Self-Evolution Framework for Uncovering and Mitigating Implicit Domain Risks in LLMs

DGX agent

arXiv:2511.07107v3 Announce Type: replace Abstract: Ensuring the safety of Large Language Models (LLMs) is critical for real-world deployment. However, current safety measures often fail to address im

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments

DGX agent

arXiv:2606.04171v1 Announce Type: cross Abstract: File-type classification underlies many workflows like malware triage, forensic carving, packet inspection, and storage indexing. Learned systems such

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment

DGX agent

arXiv:2606.04569v1 Announce Type: new Abstract: Underground mines present extreme conditions for autonomous robot navigation: GPS is denied, lighting is degraded, and tunnel topology is loop-rich and

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

Most AI pipelines are only as good as the data we provide them with, and that usually means PDFs or other unstructured documents. Contracts,…

DGX agent

Most AI pipelines are only as good as the data we provide them with, and that usually means PDFs or other unstructured documents. Contracts, invoices, reports... All have special layout, language, and

model-releasesjerry-liu--x
4 Jun 2026
← Previous
1…829830831832833…1326
Next →