AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,874 results
2 Jun 2026

CEAR: Certified Ensemble Adversarial Robustness in DNNs

SafetyDGX agent

arXiv:2606.01437v1 Announce Type: cross Abstract: Deep Neural Networks (DNNs) are highly susceptible to adversarial perturbations, leading to extensive research on robustness for safety-critical appli

Code2Math: Can Your Code Agent Effectively Evolve Math Problems Through Exploration?

AgentsDGX agent

arXiv:2603.03202v3 Announce Type: replace Abstract: As large language models (LLMs) advance their mathematical capabilities toward the IMO and research level, the scarcity of challenging, high-quality

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs

Model ReleasesDGX agent

arXiv:2606.01879v1 Announce Type: new Abstract: Existing research largely reduces cultural intelligence in LLMs to a knowledge-level problem, overlooking whether models can effectively utilize their a

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Generative AI and Digital Ecosystem Resilience: A Proactive Lifecycle-Based Survey

AgentsDGX agent

arXiv:2606.00136v1 Announce Type: cross Abstract: The proliferation of adversarial synthetic content, accelerated by Generative AI (GenAI) is rendering traditional reactive detection methods ineffecti

Introducing workspaces for Lambda Cloud

HardwareDGX agent

Lambda workspaces help teams organize cloud resources, control access, and separate dev, staging, and production in shared GPU environments. A junior researcher kills a production training run. A cont

Microsoft releases ASSERT, an open-source framework that lets developers generate and run AI behavior tests using natural-language descriptions (Ram Iyer/TechCrunch)

SafetyDGX agent

Ram Iyer / TechCrunch: Microsoft releases ASSERT, an open-source framework that lets developers generate and run AI behavior tests using natural-language descriptions — AI researchers and labs have ad

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants

SafetyDGX agent

arXiv:2508.08337v3 Announce Type: replace-cross Abstract: Algorithmic fairness research has largely framed unfairness as discrimination along sensitive attributes. However, this approach limits visibi

Pramana: Fine-Tuning Large Language Models for Epistemic Reasoning through Navya-Nyaya

Model ReleasesDGX agent

arXiv:2604.04937v1 Announce Type: cross Abstract: Large language models produce fluent text but struggle with systematic reasoning, often hallucinating confident but unfounded claims. When Apple resea

Richer Representations for Neural Algorithmic Reasoning via Auxiliary Reconstruction

TutorialsDGX agent

arXiv:2606.00559v1 Announce Type: cross Abstract: Neural algorithmic reasoning has emerged as a popular research direction. It aims to train neural networks to mimic the step-by-step behavior of class

Ryze: Evidence-Enriched Data Synthesis from Biomedical Papers

Model ReleasesDGX agent

arXiv:2606.00902v1 Announce Type: new Abstract: General-purpose VLMs remain unreliable for biomedical research because valid answers in scientific papers depend on evidence split across figures, table

Snowflake moves up the AI stack – but the System of Intelligence is still being built

IndustryDGX agent

This research note is based on four primary inputs: 1) An assessment of Snowflake Inc.’s announcements at this year’s Summit; 2) Information captured in private analyst and journalist sessions with Sn

The Case for Model Science: Verify, Explore, Steer, Refine

Model ReleasesDGX agent

arXiv:2606.01189v1 Announce Type: new Abstract: We argue that the AI community is now ready to move beyond benchmarking and consolidate scattered efforts in model analysis into a systematic discipline

The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete

Model ReleasesDGX agent

arXiv:2606.00048v1 Announce Type: cross Abstract: Prior research has established that instruction-tuned large language models exhibit left-of-center political bias, measured exclusively through abstra

ToolFG: Towards Well-Grounded Fine-Grained Image Classification

SafetyDGX agent

arXiv:2606.02518v1 Announce Type: new Abstract: Fine-grained image classification (FGIC) has broad applications and has attracted significant research attention. In this paper, we explore a novel para

Verification is the hidden bottleneck for knowledge work agents, especially in legal AI — complex, long-horizon work is graded by rubrics wi…

TutorialsDGX agent

Verification is the hidden bottleneck for knowledge work agents, especially in legal AI — complex, long-horizon work is graded by rubrics with dozens of strict criteria. In new research with @LangChai

1 Jun 2026

AI Loss of Control Incident Management: Response & Resilience

SafetyDGX agent

arXiv:2605.30406v1 Announce Type: cross Abstract: Recent research demonstrating AI systems exhibiting deception and shutdown resistance suggests that AI loss of control (LOC) is an urgent policy conce

BIAS-ID: A Framework for Analyzing Transformation Biases in AI-Generated Image Detectors

SafetyDGX agent

arXiv:2605.31153v1 Announce Type: new Abstract: Given the surge of harmful AI-generated imagery online, reliably distinguishing authentic images from generated ones has become an urgent research topic

Calibrated Uncertainty for Trustworthy Clinical Gait Analysis Using Probabilistic Multiview Markerless Motion Capture

SafetyDGX agent

arXiv:2601.22412v2 Announce Type: replace Abstract: Video-based human movement analysis holds potential for movement assessment in clinical practice and research. However, the clinical implementation

Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery

SafetyDGX agent

arXiv:2509.20941v2 Announce Type: replace Abstract: As surgical AI transitions from pixel-level detection to complex reasoning, Scene Graphs (SGs) offer the structured, relational representations nece

If LLMs Have Human-Like Attributes, Then So Does Age of Empires II

AgentsDGX agent

arXiv:2605.31514v1 Announce Type: cross Abstract: Much research has been carried out on large language models (LLMs) and LLM-powered agentic workflows. However, many works within the field state emerg

Knowledge Boundary Probing and Demand-Guided Intervention for LLM-Based Power System Code Generation

Model ReleasesDGX agent

arXiv:2605.31478v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to automate power-system analysis, but many utilities and energy-research labs require on-premise s

KnowledgeGain: Evaluating and Optimizing Science News Generation for Reader Learning

TutorialsDGX agent

arXiv:2605.31099v1 Announce Type: cross Abstract: Science news is an important medium to communicate discoveries between the research communities and the public. Yet, most metrics for generated or sum

Nvidia announcing a 550B model wasn't on my bingo card They are now the strongest american open-source lab

HardwareDGX agent

Nvidia announced development of a 550 billion parameter language model, positioning itself as a leading open-source AI research organization competing with traditional academic and independent labs. T

On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets

Model ReleasesDGX agent

arXiv:2605.31142v1 Announce Type: cross Abstract: Large-scale multilingual text embedding models play crucial role in both research and industry, yet their behavior in language-specific, multi-task se

ReTabAD: A Benchmark for Restoring Semantic Context in Tabular Anomaly Detection

Model ReleasesDGX agent

arXiv:2510.02060v2 Announce Type: replace Abstract: In tabular anomaly detection (AD), textual semantics often carry critical signals, as the definition of an anomaly is closely tied to domain-specifi

Today we’re announcing London as Runway’s European HQ, and doubling down on our investment in the UK AI ecosystem. In under two years, our L…

HardwareDGX agent

Today we’re announcing London as Runway’s European HQ, and doubling down on our investment in the UK AI ecosystem. In under two years, our London team has become central to our research, including fou

Towards Effective Long-Video Event Prediction via Multi-Level Event Semantics Mining

ApplicationsDGX agent

arXiv:2605.31069v1 Announce Type: cross Abstract: Accurately predicting future events is fundamental to content understanding and decision-making across various domains. While prior research has prima

Your Multimodal Speech Model Says I Have a Face for Radio

Model ReleasesDGX agent

arXiv:2605.30472v1 Announce Type: new Abstract: As large neural models have become better at language tasks, researchers are increasingly building multi- and omnimodal models that handle more modaliti

31 May 2026

They call it stupid hot for a reason: Heat muddles animal brains

TutorialsDGX agent

Research shows that animals get their minds muddled during heat waves, with birds struggling to learn, dogs biting more often, and chamois picking fights when temperatures rise. When animals can't sta

What’s new in Microsoft Foundry | May 2026

Model ReleasesDGX agent

May ships trace-based evaluation for any agent on any cloud, Grok 4.3 and DeepSeek V4 in the model catalog, GPT-5 Reinforcement Fine-Tuning at gated GA, three Microsoft Research on-device agent models

30 May 2026

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the o…

SafetyDGX agent

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the open on @huggingface, so researchers everywhere can scrutiniz

Increasingly, HTML Artifacts are becoming a core part of how I work with AI agents. Long-horizon agent sessions need a better way to surface…

AgentsDGX agent

Increasingly, HTML Artifacts are becoming a core part of how I work with AI agents. Long-horizon agent sessions need a better way to surface insights about what work it has done. This may not be obvio

Running Python ASGI apps in the browser via Pyodide + a service worker

Model ReleasesDGX agent

Research: Running Python ASGI apps in the browser via Pyodide + a service worker Datasette Lite is my version of Datasette that runs entirely in the browser using Pyodide in WebAssembly. When I first

29 May 2026

Error as a Lens: Probing LLM Reasoning through Synthetic Misconception Generation

AgentsDGX agent

arXiv:2605.29007v1 Announce Type: new Abstract: Personalized tutoring, teacher training, and education research need access to targeted synthetic misconceptions, but privacy and IRB constraints make l

F-RNG: Feed-Forward Relightable Neural Gaussians

ApplicationsDGX agent

arXiv:2605.25975v2 Announce Type: replace-cross Abstract: Capturing relightable 3D assets from real-world objects is a widely researched problem. Several per-scene optimization-based methods, based on

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too m…

SafetyDGX agent

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too much time in/around the big AI Labs (research for the book he

I've been using state-of-the-art models to teach small models running on my computer how I work. The result : a personal agent that runs my …

AgentsDGX agent

I've been using state-of-the-art models to teach small models running on my computer how I work. The result : a personal agent that runs my inbox, my deal pipeline, my blog, my calendar, & my research

Jailbreaking and Mitigation of Vulnerabilities in Large Language Models

SafetyDGX agent

arXiv:2410.15236v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence by advancing natural language understanding and generation, enabling app

LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning

Model ReleasesDGX agent

arXiv:2605.29649v1 Announce Type: new Abstract: Heuristic search is the dominant paradigm in symbolic AI planning, and the strongest heuristics are the result of decades of work by planning researcher

MATANet: A Multi-context Attention and Taxonomy-Aware Network for Fine-Grained Underwater Recognition of Marine Species

SafetyDGX agent

arXiv:2601.03729v2 Announce Type: replace Abstract: Fine-grained recognition of marine organisms is important for ecological research, biodiversity monitoring, habitat conservation, and evidence-based

MEMENTO: Leveraging Web as a Learning Signal for Low-Data Domains

TutorialsDGX agent

arXiv:2605.29795v1 Announce Type: new Abstract: Real-world tasks often lack large labeled datasets, motivating extensive work on learning in low-data regimes. However, existing approaches such as few-

OpenJarvis: a local-first personal AI is now available to run with Ollama Built by Stanford’s @HazyResearch and Scaling Intelligence labs, a…

Local AiDGX agent

OpenJarvis: a local-first personal AI is now available to run with Ollama Built by Stanford’s @HazyResearch and Scaling Intelligence labs, as part of their “Intelligence Per Watt” research into effici

Physics Is All You Need? A Case Study in Physicist-Supervised AI Development of Scientific Software

Model ReleasesDGX agent

arXiv:2605.30353v1 Announce Type: new Abstract: Are AI agents tools, co-authors, or researchers? We present a quantified case study (N=1): a physicist supervising an AI coding agent (Claude Code, Sonn

ProjectionBench: Evaluating Scientific Hypothesis Generation in LLMs Under Progressive Information Disclosure

Model ReleasesDGX agent

arXiv:2605.30284v1 Announce Type: new Abstract: Scientific discovery is an inherently creative and uncertain process, requiring reasoning beyond the recall of known knowledge. While many benchmarks ha

Specialty-Specific Medical Language Model for Immune-Mediated Diseases

ApplicationsDGX agent

arXiv:2605.28838v1 Announce Type: cross Abstract: Extracting detailed clinical information from free-text medical narratives remains a practical challenge for researchers and healthcare systems. Termi

Towards Understanding the Shape of Representations in Protein Language Models

Local AiDGX agent

arXiv:2509.24895v2 Announce Type: replace Abstract: While protein language models (PLMs) are one of the most promising avenues of research for future de novo protein design, the way in which they tran

28 May 2026

AI in the Workplace: The Impact of AI on Perceived Job Decency and Meaningfulness

ApplicationsDGX agent

arXiv:2605.28680v1 Announce Type: cross Abstract: The proliferation of Artificial Intelligence (AI) in workplaces is transforming how we work. While existing research on human-AI collaboration at work

Cyberbullying Governance on Social Media: A Unified Framework from Content Identification to Intervention

SafetyDGX agent

arXiv:2605.27584v1 Announce Type: new Abstract: The proliferation of social media platforms and online communities has inadvertently catalyzed the spread of cyberbullying, hate speech, and other forms

DiagramRAG: A Lightweight Framework to Retrieve Scientific Diagram for Figure Generation

TutorialsDGX agent

arXiv:2605.27931v1 Announce Type: new Abstract: Scientific diagrams are essential for communicating complex methodologies in academic papers. A natural way for researchers to specify such diagrams is

DRTriton: Large-Scale Synthetic Data Driven Reinforcement Learning for Triton Kernel Generation

Model ReleasesDGX agent

arXiv:2603.21465v2 Announce Type: replace Abstract: Developing efficient CUDA kernels is a fundamental yet challenging task in the generative AI industry. Recent research leverages Large Language Mode

From Affect to Complex Behavior: Advancing Multimodal Human-Centered AI at the 10th ABAW Workshop & Competition

SafetyDGX agent

arXiv:2605.27451v1 Announce Type: new Abstract: The 10th Affective & Behavior Analysis in-the-Wild (ABAW) Workshop and Competition, held at CVPR 2026, continues to advance research on modelling, analy

Informing AI Policy Assessment using Large-Scale Simulation of Interventions

SafetyDGX agent

arXiv:2605.27395v1 Announce Type: cross Abstract: As the rapid proliferation of AI systems and harms spurs efforts in AI governance around the world, prioritizing among competing policy options has be

OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

Model ReleasesDGX agent

arXiv:2605.28158v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used to assist with operations research (OR) modeling, yet existing OR-oriented benchmarks often redu

Performance and Explainability Requirements of Evolutionary Algorithms in Real-World Physics-Informed Optimization

ApplicationsDGX agent

arXiv:2605.28164v1 Announce Type: cross Abstract: Evolutionary computation offers a variety of tools to solve complex real-world optimization problems. However, research often focuses on smaller, simp

SA4Depth: Consistent Pose-Depth Scale Alignment for Self-Supervised Monocular Depth Estimation

SafetyDGX agent

arXiv:2605.28477v1 Announce Type: new Abstract: Self-supervised depth estimation from monocular sequences relies on the joint learning of a depth and a pose network. Despite abundant research done to

Show, Don't TELL: Explainable AI-Generated Text Detection

ApplicationsDGX agent

arXiv:2605.27921v1 Announce Type: new Abstract: Research on AI-generated text detection has presented a number of approaches to discern human from AI prose, some of which achieving high in-distributio

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

Model ReleasesDGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

Trust Me, I'm an Expert: Decoding and Steering Authority Bias in Large Language Models

SafetyDGX agent

arXiv:2601.13433v3 Announce Type: replace Abstract: Prior research demonstrates that performance of language models on reasoning tasks can be influenced by suggestions, hints and endorsements. However

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequo…

Model ReleasesDGX agent

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequoia. This investment will help us advance our research and expa

← Previous
1…354355356357358…432
Next →