AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
6 Jun 2026

GenTI: Benchmarking LLMs for Autonomous IDPS Rule Generation for Unseen Attacks

Model ReleasesDGX agent

arXiv:2606.05844v1 Announce Type: cross Abstract: Rule-based Intrusion Detection and Prevention Systems (IDPS) offer precise attack detection as well as mitigation, however their manually crafted, sig

Goedel-Architect: Streamlining Formal Theorem Proving with Blueprint Generation and Refinement

Model ReleasesDGX agent

arXiv:2606.06468v1 Announce Type: new Abstract: We introduce Goedel-Architect, an agentic framework for formal theorem proving in Lean 4 centered on blueprint generation and refinement. A blueprint is

I built a small Windows tool to monitor and manage Ollama more easily

Local AiDGX agent

A tiny Windows system tray tool that monitors local Ollama runtime with quick visual feedback about status, resource usage, and models . The app uses color-coded tray icons for quick status checks and

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

I want AI to succeed, and have been positive about many systems (Claude Code, AlphaFold, AlphaGeometry, Cicero, etc). It’s not AI that I hat…

Model ReleasesDGX agent

I want AI to succeed, and have been positive about many systems (Claude Code, AlphaFold, AlphaGeometry, Cicero, etc). It’s not AI that I hate; it’s bullshit and greed that I despise. Not my fault that

Introducing Harness-1, a 20B search agent trained with a state-externalizing harness. > frontier-level long-horizon search, rivaling Opus-4.…

Model ReleasesDGX agent

Introducing Harness-1, a 20B search agent trained with a state-externalizing harness. > frontier-level long-horizon search, rivaling Opus-4.6 and outperforming GPT-5.4 > Context-1-level cost and laten

it’s a bad day when @scaling01 goes all Gary Marcus on the media 🤣

Model ReleasesDGX agent

it’s a bad day when @scaling01 goes all Gary Marcus on the media 🤣 This is a good example of how AI news gets distorted. The example is from Tagesschau, the most watched news program in Germany. Anthr

“I've tried to find where this stops being circular and I can't.”

Model ReleasesDGX agent

“I've tried to find where this stops being circular and I can't.” 🦔Google signed a deal to pay SpaceX 920 million a month for access to 110,000 Nvidia GPUs at SpaceX data centers. The contract runs Oc

LLM Self-Recognition: Steering and Retrieving Activation Signatures

ResearchDGX agent

arXiv:2606.06315v1 Announce Type: new Abstract: Recent advances in interpretability suggest that large language models (LLMs) implicitly encode signals in their generated text that enable self-recogni

Memory is Reconstructed, Not Retrieved: Graph Memory for LLM Agents

Model ReleasesDGX agent

arXiv:2606.06036v1 Announce Type: new Abstract: Despite recent progress, LLM agents still struggle with reasoning over long interaction histories. While current memory-augmented agents rely on a stati

OPRD: On-Policy Representation Distillation

Model ReleasesDGX agent

arXiv:2606.06021v1 Announce Type: cross Abstract: On-policy distillation (OPD) supervises the student only in output space by matching next-token probabilities. This output-only paradigm has two limit

PC Layer: Polynomial Weight Preconditioning for Improving LLM Pre-Training

Model ReleasesDGX agent

arXiv:2606.06470v1 Announce Type: cross Abstract: We propose a preconditioning (PC) layer, a weight parameterization via polynomial preconditioner that ensures stable weight conditioning throughout LL

ProfiliTable: Profiling-Driven Tabular Data Processing via Agentic Workflows

Model ReleasesDGX agent

arXiv:2605.12376v2 Announce Type: replace Abstract: Table processing-including cleaning, transformation, augmentation, and matching-is a foundational yet error-prone stage in real-world data pipelines

PSEBench: A Controllable and Verifiable Benchmark for Evaluating LLMs in Patient Safety Event Triage

Model ReleasesDGX agent

arXiv:2606.05463v1 Announce Type: new Abstract: Patient safety event triage, determining whether a clinical event is reportable under jurisdiction-specific policy, is a high-stakes task typically perf

Reformulating Neural Operators in d+1 Dimensions for Embedding Evolution

ResearchDGX agent

arXiv:2505.11766v4 Announce Type: replace-cross Abstract: Neural Operators (NOs) are powerful architectures for learning mappings between function spaces. While most advances focus on refining kernel

Retry Policy Gradients in Continuous Action Spaces

Model ReleasesDGX agent

arXiv:2606.05888v1 Announce Type: new Abstract: Retry-based objectives such as pass@K and max@K optimize the best return obtained from multiple sampled trajectories, and recent work has shown that the

Reward Learning through Ranking Mean Squared Error

Model ReleasesDGX agent

arXiv:2601.09236v3 Announce Type: replace-cross Abstract: Reward design remains a significant bottleneck in applying reinforcement learning (RL) to real-world problems. A popular alternative is reward

SAGE: Scalable AI Governance & Evaluation

SafetyDGX agent

arXiv:2602.07840v3 Announce Type: replace-cross Abstract: Evaluating relevance in large-scale search systems is fundamentally constrained by the governance gap between nuanced, resource-constrained hu

SciVisAgentSkills: Design and Evaluation of Agent Skills for Scientific Data Analysis and Visualization

Model ReleasesDGX agent

arXiv:2606.05525v1 Announce Type: new Abstract: Recent advances in agentic visualization have enabled the translation of natural language into executable scientific visualization (SciVis) workflows. W

Search-Time Contamination in Deep Research Agents: Measuring Performance Inflation in Public Benchmark Evaluation

Model ReleasesDGX agent

arXiv:2606.05241v1 Announce Type: cross Abstract: Public benchmarks enable fair and reproducible evaluation of LLM reasoning, but they become fragile for deep research agents that actively search the

⚠️⚠️ Seismic shift ⚠️⚠️ It’s a good day to be Mistral. Nobody is going to trust an American AI company that is partly owned by the US Govern…

Model ReleasesDGX agent

⚠️⚠️ Seismic shift ⚠️⚠️ It’s a good day to be Mistral. Nobody is going to trust an American AI company that is partly owned by the US Government. Just the way the US doesn’t trust Huawei. After this m

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks

SafetyDGX agent

arXiv:2606.05609v1 Announce Type: cross Abstract: As large language models (LLMs) are widely deployed, identifying their vulnerability through jailbreak attacks becomes increasingly critical. Optimiza

“.. So now Google and Anthropic pay rent on the hardware Grok couldn't use, and that rent is the AI revenue story SpaceX takes public on Thu…

Model ReleasesDGX agent

“.. So now Google and Anthropic pay rent on the hardware Grok couldn't use, and that rent is the AI revenue story SpaceX takes public on Thursday.” 🦔Google signed a deal to pay SpaceX 920 million a mo

Spent more time learning deepagents from @LangChain with primary focus on integrating MCP servers with auth that doesn't conform completely …

Model ReleasesDGX agent

Spent more time learning deepagents from @LangChain with primary focus on integrating MCP servers with auth that doesn't conform completely to the OAuth standard. ( In prep of a work coming my way ) B

Starlink launches now substantially outnumber all other satellite launch sources.

Model ReleasesDGX agent

SpaceX's Starlink satellite launches have become the dominant source of orbital launches, exceeding the combined launch activity of all other satellite operators and launch providers. This milestone r

Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces

Model ReleasesDGX agent

arXiv:2606.05464v1 Announce Type: new Abstract: Verifiable reward training has improved mathematical and coding reasoning, but these domains capture only part of step-by-step decision making. Many rea

Wordle 1,812 4/6 ⬛⬛⬛⬛⬛ 🟨⬛🟨⬛🟨 ⬛🟨🟨⬛🟨 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post shows a completed Wordle game (#1,812) solved in 4 attempts, displaying the colored tile feedback pattern (gray for incorrect letters, yellow for correct letters in wrong positions, green fo

You can’t let this happen, @DavidSacks, cc @elonmusk. It’s WWWIII but with AI. Nobody wins.

Model ReleasesDGX agent

You can’t let this happen, @DavidSacks, cc @elonmusk. It’s WWWIII but with AI. Nobody wins. ⚠️⚠️ Seismic shift ⚠️⚠️ It’s a good day to be Mistral. Nobody is going to trust an American AI company that

Zero knowledge verification for frontier AI training is possible

SafetyDGX agent

arXiv:2606.05433v1 Announce Type: new Abstract: Frontier AI governance frameworks increasingly use cumulative training compute as the primary criterion for designating high-impact models, but enforcem

5 Jun 2026

29 more Starlink satellites. Over 10,000 in orbit now.

Model ReleasesDGX agent

SpaceX launched 29 additional Starlink satellites, bringing the total constellation to over 10,000 satellites in orbit. This milestone represents a significant expansion of the Starlink mega-constella

Adaptive Tokenisation Via Temporal Redundancy Masking And Latent Inpainting

Model ReleasesDGX agent

arXiv:2606.06158v1 Announce Type: new Abstract: Adaptive video tokenisation seeks to dynamically allocate token budgets based on the underlying visual complexity of a sequence. Current continuous-regi

Agents' Last Exam

Model ReleasesDGX agent

arXiv:2606.05405v1 Announce Type: cross Abstract: Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deploym

// Agents' Last Exam // Agents' Last Exam is a living benchmark of over 1,000 economically valuable tasks, built with 250+ industry experts …

Model ReleasesDGX agent

// Agents' Last Exam // Agents' Last Exam is a living benchmark of over 1,000 economically valuable tasks, built with 250+ industry experts and mapped to the U.S. federal occupational taxonomy. The ha

An issue caused some user accounts to be incorrectly suspended. We’re restoring access and working through related subscription and credit i…

Model ReleasesDGX agent

OpenAI experienced a technical issue that resulted in some user accounts being incorrectly suspended. The company announced it is working to restore access to affected accounts and address related iss

Ask or Assume? Uncertainty-Aware Clarification-Seeking in Coding Agents

AgentsDGX agent

arXiv:2603.26233v2 Announce Type: replace Abstract: As Large Language Model (LLM) agents are increasingly deployed in open-ended domains like software engineering, they frequently encounter underspeci

AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents

Model ReleasesDGX agent

arXiv:2606.05557v1 Announce Type: new Abstract: A situated query like 'where is Lin Wei?' often encodes more than its literal content: the user may also want to know whether Lin Wei is free, in a good

b9529

Local AiDGX agent

b9529 is a release tag for llama.cpp, a C/C++ implementation of LLM inference that enables running large language models locally on consumer hardware. This specific release represents a particular ver

BRepCLIP: Contrastive Multimodal Pretraining on BRep Primitives for CAD Understanding

ResearchDGX agent

arXiv:2606.05515v1 Announce Type: new Abstract: Learning representations of CAD models is a largely open problem. While 3D representation learning has flourished around point clouds and meshes, the na

CHALIS: A Challenge Dataset for Language Identification in Difficult Scenarios

Model ReleasesDGX agent

arXiv:2606.06088v1 Announce Type: new Abstract: We present CHALIS (Challenging Language Identification Samples), a new benchmark dataset explicitly designed to address difficult cases in language iden

CollabSim: A CSCW-Grounded Methodology for Investigating Collaborative Competence of LLM Agents through Controlled Multi-Agent Experiments

AgentsDGX agent

arXiv:2606.06399v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models have shown growing promise, with their effectiveness resting on agents' ability to coordinate t

Critical context on the new Anthropic blog: 1, AGI is *harder* than RSI (as used below). AGI: machine can do anything human can do, autonomo…

Model ReleasesDGX agent

Critical context on the new Anthropic blog: 1, AGI is *harder* than RSI (as used below). AGI: machine can do anything human can do, autonomously [not achieved] RSI (as used below): AI is a useful codi

Efficient Punctuation Restoration via Weighted Lookahead Scoring Method for Streaming ASR Systems

Model ReleasesDGX agent

arXiv:2606.05179v1 Announce Type: new Abstract: Punctuation restoration improves ASR (Automatic Speech Recognition) readability. However streaming ASR requires online decisions with limited future con

EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026

Local AiDGX agent

arXiv:2605.24496v2 Announce Type: replace Abstract: The EPIC-KITCHENS-100 Action Detection challenge evaluates whether a model can localize the start and end of each action in long untrimmed egocentri

even though @activegraphai is not a memory tool, it's a runtime built around memory so it can do well there cool to see a third party verify…

Model ReleasesDGX agent

even though @activegraphai is not a memory tool, it's a runtime built around memory so it can do well there cool to see a third party verify this and see that activegraph does well on privacy programs

FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization

SafetyDGX agent

arXiv:2606.05468v1 Announce Type: new Abstract: Post-training Vision-Language-Action (VLA) models into policies that can be reliably deployed on real robots remains a major bottleneck. SFT and DAgger

Gemma 4 QAT is here. Available for all sizes of Gemma 4, optimized with Quantization-Aware Training (QAT) to reduce memory requirements whil…

Model ReleasesDGX agent

Gemma 4 QAT is here. Available for all sizes of Gemma 4, optimized with Quantization-Aware Training (QAT) to reduce memory requirements while preserving performance. Live now in LM Studio. https://lms

Google Cloud says its SpaceX compute deal is a 'short-term' agreement 'to ensure we have bridge capacity to meet surging customer demand' for Gemini Enterprise (Kate Conger/New York Times)

Model ReleasesDGX agent

Kate Conger / New York Times: Google Cloud says its SpaceX compute deal is a “short-term” agreement “to ensure we have bridge capacity to meet surging customer demand” for Gemini Enterprise — Elon Mus

I always appreciate the opportunity to discuss @LawZero_ and our approach to honest, reliable AI. Working on the Scientist AI with my brilli…

Model ReleasesDGX agent

I always appreciate the opportunity to discuss @LawZero_ and our approach to honest, reliable AI. Working on the Scientist AI with my brilliant colleagues at LawZero has made me very confident that we

Interpreting Style Representations via Style-Eliciting Prompts

ResearchDGX agent

arXiv:2606.05716v1 Announce Type: new Abstract: Style representation learning is a powerful tool for authorship analysis and modeling writing style, yet the latent nature of learned representations ma

Latent Implicit Visual Reasoning

ResearchDGX agent

arXiv:2512.21218v2 Announce Type: replace Abstract: While Large Multimodal Models (LMMs) have made significant progress, they remain largely text-centric, relying on language as their core reasoning m

LatentSkill: From In-Context Textual Skills to In-Weight Latent Skills for LLM Agents

Model ReleasesDGX agent

arXiv:2606.06087v1 Announce Type: new Abstract: Agent systems increasingly use textual skills to encode reusable task procedures, but injecting these skills into the prompt at every step incurs substa

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios

SafetyDGX agent

arXiv:2606.05952v1 Announce Type: new Abstract: In this work, we propose an agentic gamification framework for hazard-informed learning of robot safety policies through synthetic scenarios. We model s

Many Circuits, One Mechanism: Input Variation and Evaluation Granularity in Circuit Discovery

ResearchDGX agent

arXiv:2606.06267v1 Announce Type: new Abstract: Circuit discovery methods identify subgraphs that explain specific model behaviors, and structural differences between discovered circuits are commonly

MDP-GRPO: Stabilized Group Relative Policy Optimization for Multi-Constraint Instruction Following

Model ReleasesDGX agent

arXiv:2606.06058v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards is ideal for multi-constraint instruction following, yet standard group-relative policy optimization (G

Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense

SafetyDGX agent

arXiv:2606.05743v1 Announce Type: cross Abstract: Despite advances in safety alignment, large language models remain vulnerable to continuously evolving jailbreaks. Existing fine-tuned safety classifi

Oklch+: A Three-Parameter Extension of Oklab for Improved Color Difference Prediction

Model ReleasesDGX agent

arXiv:2606.05255v1 Announce Type: cross Abstract: Oklab and its cylindrical representation Oklch are widely adopted in interpolation and design workflows as perceptually motivated color spaces, but th

OLIVE: Online Low-Rank Incremental Learning for Efficient Adaptive Exoskeletons

Model ReleasesDGX agent

arXiv:2606.05234v1 Announce Type: new Abstract: Wearable exoskeleton systems hold promise for restoring mobility in individuals with physical impairments, yet most existing controllers rely on static

Operation-Guided Progressive Human-to-AI Text Transformation Benchmark for Multi-Granularity AI-Text Detection

Model ReleasesDGX agent

arXiv:2606.06481v1 Announce Type: new Abstract: As AI writing assistants become increasingly integrated into real-world drafting and revision workflows, many documents are no longer purely human-writt

Parallel Jacobi Decoding for Fast Autoregressive Image Generation

ResearchDGX agent

arXiv:2606.05703v1 Announce Type: new Abstract: Autoregressive (AR) models have demonstrated remarkable performance in generating high-fidelity images. However, their inherently sequential next-token

Physics-Guided Deep Unfolding for Blind Cross-Sensor Spectral Super-Resolution via Learning the Spectral Transformation Function

Model ReleasesDGX agent

arXiv:2606.05759v1 Announce Type: new Abstract: Hyperspectral imaging provides rich spectral information for quantitative remote sensing, yet hyperspectral sensors remain costly and thus unavailable i

Reinforcement Learning Elicits Contextual Learning of Unseen Language Translation

ResearchDGX agent

arXiv:2606.06428v1 Announce Type: new Abstract: Prior work has shown that large language models (LLMs) can translate unseen or low-resource languages by undergoing continued training or even by encodi

← Previous
1…647648649650651…1042
Next →