AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,022 results
Research

Hierarchies of Calibration: Classification meets Regression

DGX agent

arXiv:2606.03245v1 Announce Type: cross Abstract: Concepts of calibration formalize the compatibility between probabilistic predictions and the respective outcomes. In a nutshell, the outcomes ought t

researcharxiv-cs-lg
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also …

DGX agent

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also drop a full guide on how it works! Building an agent that ca

agentstogether-ai--x
3 Jun 2026
Model Releases

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he h…

DGX agent

Jerry Liu built one of the most installed pieces of AI plumbing of the last three years. Then he sat down and told me the framework era he helped create is over. The agent harness ate the abstraction

model-releasesjerry-liu--x
3 Jun 2026
Safety

LAP: An Agent-to-Instrument Protocol for Autonomous Science

DGX agent

arXiv:2606.03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure. Large language model agents now plan experiments, and self-driving laboratories execu

safetyarxiv-cs-ai
3 Jun 2026
Research

Learning Multi-Scale Hypergraph for High-Order Brain Connectivity Analysis

DGX agent

arXiv:2606.03310v1 Announce Type: cross Abstract: Understanding complex interactions between brain regions is critical for early neurodegenerative disease classification such as Alzheimer's Disease (A

researcharxiv-cs-ai
3 Jun 2026
Agents

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or dist…

DGX agent

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or distillation from previous models. this means reasoning, agentic

agentsswyx--x
3 Jun 2026
Model Releases

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

DGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

model-releasesarxiv-cs-lg
3 Jun 2026
Agents

MUSE: A Unified Agentic Harness for MLLMs

DGX agent

arXiv:2606.03005v1 Announce Type: cross Abstract: Despite rapid progress, multimodal large language models (MLLMs) still fail on tasks that humans solve effortlessly, such as navigating a grid maze fr

agentsarxiv-cs-ai
3 Jun 2026
Applications

Not to mention that having HIPAA and FERPA compliant AI systems makes thousands of students and researchers using them less risky.

DGX agent

HIPAA and FERPA compliant AI systems reduce institutional and legal risks when used by students and researchers by ensuring sensitive health and educational data are properly protected. Compliance wit

applicationsethan-mollick--x
3 Jun 2026
Agents

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

DGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

agentsarxiv-cs-ai
3 Jun 2026
Research

Optimizing Neuro-Fuzzy and Colonial Competition Algorithms for Skin Cancer Diagnosis in Dermatoscopic Images

DGX agent

arXiv:2505.08886v2 Announce Type: replace Abstract: The rising incidence of skin cancer, coupled with limited public awareness and a shortfall in clinical expertise, underscores an urgent need for adv

researcharxiv-cs-cv
3 Jun 2026
Model Releases

Plan2Map: A Multimodal Benchmark for Document-Grounded Geospatial Boundary Reconstruction from Planning Records

DGX agent

arXiv:2606.02747v1 Announce Type: cross Abstract: Planning records define restrictions over geographic areas, but their source documents often provide only indirect spatial evidence rather than machin

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

PyraMathBench: Evaluating and Improving Mathematical Capability in Large Language Models

DGX agent

arXiv:2606.03858v1 Announce Type: new Abstract: Despite the pivotal role of numerical reasoning as the cornerstone of mathematical capabilities in large language models (LLMs) across applications, few

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Reasoning Structure of Large Language Models

DGX agent

arXiv:2606.03883v1 Announce Type: new Abstract: Large reasoning models (LRMs) are often evaluated using metrics such as final-answer accuracy or token count. However, identical scores on these metrics

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

DGX agent

arXiv:2606.03994v1 Announce Type: new Abstract: Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image

safetyarxiv-cs-cv
3 Jun 2026
Agents

SPADE: Sketch-guided Path Planning Augmented with Diffusion Experts

DGX agent

arXiv:2606.03512v1 Announce Type: cross Abstract: Path planning is essential for Autonomous Mobile Robots (AMRs). Conventional methods for incorporating human preferences into planning typically rely

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks

DGX agent

arXiv:2606.03606v1 Announce Type: cross Abstract: Large language models achieve strong performance on arithmetic reasoning benchmarks, and one common response to arithmetic brittleness is to delegate

model-releasesarxiv-cs-ai
3 Jun 2026
Research

The Hermes Web Dashboard got a major overhaul: it is now a feature-complete admin panel that you can manage entirely from your browser.

DGX agent

The Hermes Web Dashboard has been significantly redesigned to function as a comprehensive admin panel with full feature parity, allowing users to manage all operations directly through a web browser i

researchnous-research--x
3 Jun 2026
Research

The next chapter in flood resilience: Open sourcing Google’s hydrology framework

DGX agent

Google has open-sourced its flood forecasting framework, which replicates operational FloodHub model training settings and reflects methodology described in a 2024 Nature paper for global ungauged flo

researchgoogle-research
3 Jun 2026
Research

The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Constructs

DGX agent

arXiv:2606.03357v1 Announce Type: cross Abstract: When prompting SLMs for psychometric assessments, researchers assume the outputs reflect semantic reasoning. We evaluate this premise across 13 open-w

researcharxiv-cs-ai
3 Jun 2026
Research

this is an interesting point in the new ted chiang piece – no one really claims that alphafold is conscious, or that sora or midjourney or d…

DGX agent

Yann LeCun discusses a point from a Ted Chiang piece about the lack of claims regarding consciousness in recent AI systems like AlphaFold, Sora, and Midjourney, suggesting skepticism about attributing

researchyann-lecun--x
3 Jun 2026
Model Releases

This is really intellectually dishonest. I have not been arguing that LLM token prices are increasing (though the all you can eat buffet is …

DGX agent

This is really intellectually dishonest. I have not been arguing that LLM token prices are increasing (though the all you can eat buffet is over), I have been arguing the *opposite*, viz that they wil

model-releasesgary-marcus--x
3 Jun 2026
Agents

This SkillOpt paper from Microsoft is a must-read! (bookmark it) I was a bit skeptical of the results reported in the paper when I shared it…

DGX agent

This SkillOpt paper from Microsoft is a must-read! (bookmark it) I was a bit skeptical of the results reported in the paper when I shared it a few days ago. However, I managed to integrate it into my

agentsdair-ai--x
3 Jun 2026
Safety

Towards a Science of AI Agent Reliability

DGX agent

arXiv:2602.16666v3 Announce Type: replace Abstract: AI agents are increasingly deployed to execute important tasks. While rising accuracy scores on standard benchmarks suggest rapid progress, many age

safetyarxiv-cs-ai
3 Jun 2026
Research

Tracking Urban Atmospheric Pollutants using Sentinel-5P Satellite Data

DGX agent

arXiv:2606.02592v1 Announce Type: cross Abstract: Urban nitrogen dioxide (NO_2) is a key indicator of combustion-related air pollution and exhibits strong spatial and temporal variability in cities. T

researcharxiv-cs-ai
3 Jun 2026
Model Releases

TriEval: A Resource-Efficient Pipeline for LLM Bias, Toxicity, and Truthfulness Assessment

DGX agent

arXiv:2606.03036v1 Announce Type: new Abstract: LLMs have evolved from basic chatbots to the backbone of the AI ecosystem, now widely used in healthcare, schools, and government services. The domain-w

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

TSQAgent: Rating Time Series Data Quality via Dedicated Agentic Reasoning

DGX agent

arXiv:2606.03629v1 Announce Type: new Abstract: Assessing the quality of time series (TS) data is fundamental yet inherently challenging due to the multifaceted nature of quality dimensions. Recently,

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

VistaHop: Benchmarking Multi-hop Visual Reasoning for Visual DeepSearch

DGX agent

arXiv:2606.03273v1 Announce Type: cross Abstract: Visual DeepSearch requires multimodal large reasoning model (MLRM) agents to answer complex visual queries by repeatedly inspecting image regions, gro

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning

DGX agent

arXiv:2606.02866v1 Announce Type: new Abstract: When does multi-agent debate help data cleaning, and when does it hurt? Across three benchmarks, four model families, and over 6,000 task-condition pair

agentsarxiv-cs-ai
3 Jun 2026
Research

X-RAY: Mapping LLM Reasoning Capability via Formalized and Calibrated Probes

DGX agent

arXiv:2603.05290v2 Announce Type: replace Abstract: Large language models (LLMs) achieve promising performance, yet their ability to reason remains poorly understood. Existing evaluations largely emph

researcharxiv-cs-ai
3 Jun 2026
Safety

You don’t need to be @garymarcus to know which way the wind blows.

DGX agent

You don’t need to be @garymarcus to know which way the wind blows. I would be way more bullish on AI if it actually worked and was actually replacing real humans at scale. Nothing is changing and we’r

safetygary-marcus--x
3 Jun 2026
Research

A Direct Approach for Handling Contextual Bandits with Latent State Dynamics

DGX agent

arXiv:2604.08149v2 Announce Type: replace Abstract: We consider a linear contextual bandit model where contexts and rewards are governed by a finite hidden Markov chain. We first revisit the simplifie

researcharxiv-cs-lg
2 Jun 2026
Research

A Grammar of Machine Learning Workflows: Rejecting Data Leakage at Call Time

DGX agent

arXiv:2603.10742v4 Announce Type: replace Abstract: Data leakage has been identified in 648 published papers across 30 scientific fields. The knowledge to prevent it has existed for over a decade; the

researcharxiv-cs-lg
2 Jun 2026
Safety

A Practical Upper Bound on Selection Bias Effects in Medical Prediction Models

DGX agent

arXiv:2606.00563v1 Announce Type: cross Abstract: Selection bias is a common and often unavoidable aspect of real-world data that challenges the generalizability of machine learning models. When model

safetyarxiv-cs-ai
2 Jun 2026
Research

A Sonar-Visual Dataset for Cross-Modal Underwater Robot Perception

DGX agent

arXiv:2606.01398v1 Announce Type: new Abstract: Underwater robots typically use both cameras and sonar for perception to leverage the rich semantic details of vision and the robust range measurements

researcharxiv-cs-ro
2 Jun 2026
Model Releases

A Systematic Benchmark of Intraoperative Ultrasound-to-MR Synthesis for Brain Tumour Surgery

DGX agent

arXiv:2606.00630v1 Announce Type: new Abstract: Intraoperative ultrasound (ioUS) is a versatile, cost-effective modality in brain tumour surgery, but its interpretation is difficult: acquisition plane

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Absorbing Complexity: An Interaction-Native Knowledge Harness for Financial LLM Agents

DGX agent

arXiv:2606.01886v1 Announce Type: new Abstract: Financial AI agents often fail for a simple reason: they make users carry the complexity. A user must repeatedly restate goals, risk preferences, portfo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ADRA-Bank: A Modular Benchmark for Academic Deep Research Agents

DGX agent

arXiv:2512.00986v3 Announce Type: replace Abstract: A surge in academic publications calls for automated deep research (DR) systems, but accurately evaluating them is still an open problem. First, exi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

AgentPLM: Agentic Protein Language Models with Reasoning-Augmented Decoding for Protein Sequence Design

DGX agent

arXiv:2606.02386v1 Announce Type: new Abstract: Protein language models (PLMs) are passive oracles: they generate sequences in a single forward pass with no mechanism to consult external biophysical f

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Announcing Spanner Graph algorithms: Google-grade intelligence for connected data

DGX agent

At Google Cloud Next, we announced the preview of graph algorithms with Spanner Graph, bringing Google Research’s state-of-the-art graph mining capabilities natively to your database. These graph inte

model-releasesgoogle-cloud-ai
2 Jun 2026
Agents

Application of Algorithms in Energy-Efficient Design Platforms for Green Building

DGX agent

arXiv:2606.01229v1 Announce Type: new Abstract: During green building design, computer-aided energy assessment is widely used to improve efficiency and achieve overall optimization. This paper present

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training

DGX agent

arXiv:2606.00602v1 Announce Type: new Abstract: Learning transferable and interpretable representations from medical volumetric scans remains challenging due to complex anatomical structures and weak,

model-releasesarxiv-cs-cv
2 Jun 2026
Local Ai

b9469

DGX agent

b9469 is an intermediate build release of llama.cpp, a C/C++ implementation that enables large language model inference on consumer hardware with minimal dependencies. Build releases like b9469 repres

local-aillama-cpp-releases
2 Jun 2026
Model Releases

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning

DGX agent

arXiv:2606.02109v1 Announce Type: new Abstract: Enterprise AI systems that translate natural language into SQL queries and orchestrate multi-step agentic reasoning pipelines require evaluation approac

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Bayesian Inference of Nonlinear Malaria Dynamics in Ghana via an Ensemble Markov Chain Monte Carlo Sampler

DGX agent

arXiv:2606.00783v1 Announce Type: cross Abstract: Reliable quantification of malaria dynamics in sub-Saharan Africa is hindered by short, noisy, and spatially heterogeneous surveillance records. In Gh

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

Before the Model Learns the Bug:Fuzzing RLVR Verifiers

DGX agent

arXiv:2606.01066v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) replaces human preference labels with executable reward functions such as math answer checkers, JS

tutorialsarxiv-cs-ai
2 Jun 2026
Safety

Beyond Access: Guided LLM Scaffolding for Independent Learning in Undergraduate Statistics

DGX agent

arXiv:2606.01375v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly entering students' learning practices, but their educational value depends on whether they support reaso

safetyarxiv-cs-ai
2 Jun 2026
Safety

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation

DGX agent

arXiv:2602.11790v2 Announce Type: replace Abstract: Although recent end-to-end video generation models demonstrate impressive performance in visually oriented content creation, they remain limited in

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…161162163164165…209
Next →