AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,637 results
Model Releases

If this prompt feels well written to you, it's because Suzanne is a writer in her little spare time! You can read her short story, Mall of A…

DGX agent

If this prompt feels well written to you, it's because Suzanne is a writer in her little spare time! You can read her short story, Mall of America here: https://suzannewang.com/mall-of-america It's on

model-releasesthariq--x
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

In early May, the best superforecasters predicted that, by the end of the year, the longest METR 80% task horizons would reach 3-4 hours. In…

DGX agent

In early May, the best superforecasters predicted that, by the end of the year, the longest METR 80% task horizons would reach 3-4 hours. In late May, Claude Mythos achieved that number. We also asked

model-releasesethan-mollick--x
3 Jun 2026
Safety

Jailbreak Attack Initializations as Extractors of Compliance Directions

DGX agent

arXiv:2502.09755v4 Announce Type: replace-cross Abstract: Safety-aligned LLMs respond to prompts with either compliance or refusal, each corresponding to distinct directions in the model's activation

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he h…

DGX agent

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he helped create is over. The agent harness ate the abstraction

model-releasesjerry-liu--x
3 Jun 2026
Research

Learning to Refine: Spectral-Decoupled Iterative Refinement Framework for Precipitation Nowcasting

DGX agent

arXiv:2606.02661v1 Announce Type: cross Abstract: Accurate precipitation nowcasting is vital for disaster mitigation, but deep learning methods face a key trade-off: regression models produce over-smo

researcharxiv-cs-ai
3 Jun 2026
Research

Learning to Solve, Forgetting to Retain: Correct-Set Turnover in RLVR

DGX agent

arXiv:2606.03087v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the ability of large language model, yet headline accuracy gains often conceal a hidden c

researcharxiv-cs-lg
3 Jun 2026
Tutorials

Lethe: Adapter-Augmented Dual-Stream Update for Persistent Knowledge Erasure in Federated Unlearning

DGX agent

arXiv:2601.22601v2 Announce Type: replace Abstract: Federated unlearning (FU) aims to erase designated client-level, class-level, or sample-level knowledge from a global model. Existing studies common

tutorialsarxiv-cs-lg
3 Jun 2026
Safety

Letting Tutor Personas Speak Up for LLMs: Learning Steering Vectors from Dialogue via Preference Optimization

DGX agent

arXiv:2602.07639v2 Announce Type: replace Abstract: With the emergence of large language models (LLMs) as a powerful class of generative artificial intelligence (AI), their use in tutoring has become

safetyarxiv-cs-cl
3 Jun 2026
Model Releases

Locality Does Not Imply Reachability: Boundary Repair in Block-Sparse Causal Attention

DGX agent

arXiv:2606.02680v1 Announce Type: new Abstract: Sparse causal attention is usually described by sequence locality: nearby tokens should remain easy to access, while distant tokens may be dropped to re

model-releasesarxiv-cs-lg
3 Jun 2026
Tools

M3 brings sparse attention + 1M context + multimodality, and Together did the hard serving work to make it fast. Great collaboration with th…

DGX agent

Anthropic's Claude 3.5 Sonnet (M3) model features sparse attention mechanisms, 1 million token context length, and multimodal capabilities, with Together AI optimizing the serving infrastructure to en

toolstogether-ai--x
3 Jun 2026
Research

Making Brain-Computer Interfaces More Secure

DGX agent

arXiv:2606.02597v1 Announce Type: new Abstract: The development of brain-computer interfaces (BCIs) based on electroencephalograms (EEGs) has advanced significantly mainly to machine learning. Althoug

researcharxiv-cs-lg
3 Jun 2026
Agents

MariData: One-Step Unpaired Image Translation for Maritime Environments

DGX agent

arXiv:2606.03246v1 Announce Type: new Abstract: The development on robust perception systems for Maritime Autonomous Surface Ships (MASS) is heavily constrained by the scarcity of diverse training dat

agentsarxiv-cs-cv
3 Jun 2026
Model Releases

MARIO: Motion-Augmented Real-Time Multi-Sensor Inertial Odometry

DGX agent

arXiv:2606.02996v1 Announce Type: cross Abstract: Inertial odometry (IO) using only Inertial Measurement Units (IMUs) provides a lightweight solution for human motion tracking in augmented reality (AR

model-releasesarxiv-cs-cv
3 Jun 2026
Safety

Mitigating False Credit Propagation: Probabilistic Graphical Reward Aggregation for Rubric-Based Reinforcement Learning

DGX agent

arXiv:2606.03361v1 Announce Type: new Abstract: Rubric-based rewards are increasingly used for open-ended language model post-training, but criterion-level scores are often aggregated as independent u

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

Mixed-Modality Dual Face-Hair Retrieval

DGX agent

arXiv:2606.03470v1 Announce Type: new Abstract: We introduce Dual Face-Hair Retrieval (DFHR), a new mixed-modality dual-reference task in image retrieval where a query consists of a face image specify

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

MultiTurnPSB: Evaluating Multi-Turn Jailbreak Attacks an dClassifier-Based Defenses for Medical AI Safety

DGX agent

arXiv:2606.02630v1 Announce Type: cross Abstract: Patient-facing medical chatbots are commonly evaluated on single-turn prompts, yet real users push back after refusals, add urgency, and invoke author

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

My timeline seems to have people surprised that U Chicago is getting Claude, but tons of schools (including U Penn where I teach) have schoo…

DGX agent

My timeline seems to have people surprised that U Chicago is getting Claude, but tons of schools (including U Penn where I teach) have school-wide AI There are lots of things that need to be figured o

model-releasesethan-mollick--x
3 Jun 2026
Research

OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration

DGX agent

arXiv:2507.23035v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive capabilities across a wide range of applications, but demand substantial memory and comput

researcharxiv-cs-lg
3 Jun 2026
Model Releases

OpenAI ran a hiring challenge, but the top candidate was one they couldn’t hire: our autonomous research agent, Aiden. In Parameter Golf, Ai…

DGX agent

OpenAI ran a hiring challenge, but the top candidate was one they couldn’t hire: our autonomous research agent, Aiden. In Parameter Golf, Aiden ran for 22 days, and out-outperformed all 1,016 other re

model-releasesclem-delangue--x
3 Jun 2026
Model Releases

Optimal Initialization in Depth: Lyapunov Initialization and Limit Theorems for Deep Leaky ReLU Networks

DGX agent

arXiv:2602.10949v2 Announce Type: replace-cross Abstract: Effective initialization in deep networks requires an understanding of random neural networks. In this work, a rigorous probabilistic analysis

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Optimizing Explicit Unit-Distance Lower-Bound Certificates

DGX agent

arXiv:2606.03419v1 Announce Type: cross Abstract: The 2026 disproof of Erdos's unit-distance conjecture and Sawin's subsequent explicit quantitative refinement show that the maximum number u(n) of uni

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Pathway-Structured Privileged Distillation for Deployable Computational Pathology

DGX agent

arXiv:2606.02877v1 Announce Type: new Abstract: Integrating transcriptomics and histopathology can improve cancer risk modelling, yet practical use is constrained by the limited availability of RNA pr

safetyarxiv-cs-cv
3 Jun 2026
Model Releases

PerchRL: Vision-Based Agile Perching on Inclined Platforms under Rapid and Irregular Motion

DGX agent

arXiv:2606.03441v1 Announce Type: cross Abstract: Autonomous vision-based perching of quadrotors on moving inclined platforms is critical for air-ground collaboration but remains challenging due to th

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

PINNfluence: Interpreting PINNs through Influence Functions

DGX agent

arXiv:2409.08958v3 Announce Type: replace-cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful deep learning approach for solving partial differential equations (PDEs) i

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Plan2Map: A Multimodal Benchmark for Document-Grounded Geospatial Boundary Reconstruction from Planning Records

DGX agent

arXiv:2606.02747v1 Announce Type: cross Abstract: Planning records define restrictions over geographic areas, but their source documents often provide only indirect spatial evidence rather than machin

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Planning with Uncertainty: Symmetries, Policy Inference, and Solution Compression

DGX agent

arXiv:2403.19883v2 Announce Type: replace Abstract: Fully-observable non-deterministic (FOND) planning is at the core of artificial intelligence planning with uncertainty. It models uncertainty throug

safetyarxiv-cs-ai
3 Jun 2026
Research

Position: Adversarial ML for LLMs Is Not Making Any Progress

DGX agent

arXiv:2502.02260v2 Announce Type: replace Abstract: In the past decade, considerable research effort has been devoted to securing machine learning (ML) models that operate in adversarial settings. Yet

researcharxiv-cs-lg
3 Jun 2026
Industry

Powering the Inference Era: Inside the DigitalOcean Data & Learning Layer

DGX agent

DigitalOcean discusses infrastructure and platform capabilities designed to support the inference phase of machine learning, where trained models are deployed to make predictions on new data at scale.

industrydigitalocean
3 Jun 2026
Model Releases

ProtocolBench: Which LLM MultiAgent Protocol to Choose?

DGX agent

arXiv:2510.17149v3 Announce Type: replace Abstract: As large-scale multi-agent systems evolve, the communication protocol layer has become a critical yet under-evaluated factor shaping performance and

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations

DGX agent

arXiv:2606.03136v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks on large language models (LLMs) reveal a mismatch in current guardrails: they operate on individual turns, while attacks

safetyarxiv-cs-cl
3 Jun 2026
Agents

Reproducibility is the New Copyleft: Defining AGI-oriented Reproducible Builds

DGX agent

arXiv:2606.03019v1 Announce Type: cross Abstract: Copyleft, as implemented in licenses such as the GNU General Public License, was a legal hack that used copyright to guarantee user freedom by tying t

agentsarxiv-cs-ai
3 Jun 2026
Research

Rethinking the Idiomaticity Decomposability Hypothesis: Evidence from Distributional Learning

DGX agent

arXiv:2606.03817v1 Announce Type: new Abstract: Idioms can be analysed in terms of their decomposability, the extent to which constituent meanings contribute to the figurative whole. Decomposability i

researcharxiv-cs-cl
3 Jun 2026
Safety

Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning

DGX agent

arXiv:2606.03234v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has become the dominant approach for improving mathematical reasoning in large language models, ye

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

ROBUST-WT: Robust Uncertainty-aware Segmentation Transform via Whitening and Training Enhancements

DGX agent

arXiv:2606.03069v1 Announce Type: cross Abstract: Generalized segmentation of medical images prevents performance degradation when different imaging devices and clinical protocols are used across mult

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Scalable Uncertainty Quantification for Extreme Weather Forecasting via Empirical Neural Tangent Kernels

DGX agent

arXiv:2606.02886v1 Announce Type: cross Abstract: Deep learning weather models now match numerical weather prediction accuracy while running orders of magnitude faster, but produce deterministic forec

researcharxiv-cs-ai
3 Jun 2026
Model Releases

ScoreStop: Gradient-based early stopping using functional score tests

DGX agent

arXiv:2606.02740v1 Announce Type: cross Abstract: Gradient boosted decision trees require a stopping rule to avoid overfitting. The standard rule monitors a validation loss and stops if the loss fails

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

shower thought If: 1. AI is smarter than humans at law, therapy, etc. 2. Humans still like talking to other humans. Then: Humans are just an…

DGX agent

shower thought If: 1. AI is smarter than humans at law, therapy, etc. 2. Humans still like talking to other humans. Then: Humans are just an AI wrapper. Everyone should just regurgitate what Claude te

model-releasesjerry-liu--x
3 Jun 2026
Model Releases

Sources: Benchmark raised 2B across two new funds, including a 1.25B fund for late-stage bets, its first growth fund after decades focusing on new startups (Kate Clark/Wall Street Journal)

DGX agent

Kate Clark / Wall Street Journal: Sources: Benchmark raised 2B across two new funds, including a 1.25B fund for late-stage bets, its first growth fund after decades focusing on new startups — After a

model-releasestechmeme
3 Jun 2026
Model Releases

Sources: DeepSeek is set to raise ~7.4B in its first funding round from investors including Tencent and CATL at a valuation of between ~52B and ~$59B (Reuters)

DGX agent

Reuters: Sources: DeepSeek is set to raise ~7.4B in its first funding round from investors including Tencent and CATL at a valuation of between ~52B and ~59B — Chinese AI startup DeepSeek is set to ra

model-releasestechmeme
3 Jun 2026
Local Ai

Tailoring Strictly Proper Scoring Rules for Downstream Tasks: An Application to Causal Inference

DGX agent

arXiv:2606.03332v1 Announce Type: new Abstract: Probabilistic models are typically trained using task-agnostic objectives like log-loss, which can lead to significant errors in downstream estimation.

local-aiarxiv-cs-lg
3 Jun 2026
Safety

Temporal Action Selection for Action Chunking

DGX agent

arXiv:2511.04421v2 Announce Type: replace Abstract: Action chunking is a widely adopted approach in Learning from Demonstration (LfD). By modeling multi-step action chunks rather than single-step acti

safetyarxiv-cs-ro
3 Jun 2026
Model Releases

TeX-1500: A Paired Real-World LWIR Hyperspectral Dataset and Benchmark for Temperature-Emissivity-Texture Decomposition

DGX agent

arXiv:2606.03806v1 Announce Type: new Abstract: Temperature-emissivity-texture (TeX) decomposition seeks to recover object heat state, material spectral response, and visible-like geometric texture fr

model-releasesarxiv-cs-cv
3 Jun 2026
Research

Text-attributed Graph Condensation via Text Selection and Attribute Matching

DGX agent

arXiv:2606.03839v1 Announce Type: new Abstract: Text-Attributed Graph (TAG) is an important type of graph structured data, where each node has a text description. TAG models usually train a Graph Neur

researcharxiv-cs-lg
3 Jun 2026
Model Releases

The Geometry of LLM-as-Judge: Why Inter-LLM Consensus Is Not Human Alignment

DGX agent

arXiv:2606.03043v1 Announce Type: new Abstract: LMs-as-judges are now standard, yet judges agree strongly with one another while agreeing only weakly with humans. We test whether this reflects shared

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

The Impact of Configuring Agentic AI Coding Tools on Build-vs-Buy Decisions: A Study Protocol

DGX agent

arXiv:2606.03907v1 Announce Type: cross Abstract: Agentic AI coding tools write code with increasing autonomy and in doing so decide when to import a library and when to implement functionality from s

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

The Ringelmann Effect in Multi-Agent LLM Systems: A Scaling Law for Effective Team Size

DGX agent

arXiv:2606.02646v1 Announce Type: cross Abstract: Inference-time multi-agent LLM scaling lacks a shared unit: counting nominal agents conflates cost with independent evidence. We derive a two-paramete

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMs

DGX agent

arXiv:2606.03092v1 Announce Type: new Abstract: Inference-time scaling has emerged as a critical avenue for enhancing Large Language Models' performance, yet real-world deployment is constrained by st

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

This is really intellectually dishonest. I have not been arguing that LLM token prices are increasing (though the all you can eat buffet is …

DGX agent

This is really intellectually dishonest. I have not been arguing that LLM token prices are increasing (though the all you can eat buffet is over), I have been arguing the *opposite*, viz that they wil

model-releasesgary-marcus--x
3 Jun 2026
← Previous
1…832833834835836…1326
Next →