AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,060 results
Agents

other terms i’m learning: blackboard, DDD, CQRS, Akka, Kafka, BASE…

DGX agent

This post discusses various software architecture and distributed systems concepts including the Blackboard pattern, Domain-Driven Design (DDD), Command Query Responsibility Segregation (CQRS), and me

agentsyohei-nakajima--x
28 May 2026
Research
X Post
Paper
YouTube
Reddit
GitHub

Our team at @AIatMeta is excited to announce ATLAS: one of the largest automated formalization efforts to date. ATLAS contains Lean 4 formal…

DGX agent

Our team at @AIatMeta is excited to announce ATLAS: one of the largest automated formalization efforts to date. ATLAS contains Lean 4 formalizations of both statements and proofs from 25+ mathematics

researchyann-lecun--x
28 May 2026
Agents

Out of Sight, Not Out of Mind: Unveiling Latent Attack in Latent-based Multi-Agent Systems

DGX agent

arXiv:2605.28214v1 Announce Type: cross Abstract: Latent-based multi-agent systems replace parts of explicit inter-agent communication with hidden representations, offering a new direction for efficie

agentsarxiv-cs-lg
28 May 2026
Local Ai

Outer-Momentum Restarting in High-Dimensional Two-Phase Optimization

DGX agent

arXiv:2605.28585v1 Announce Type: new Abstract: Communication-efficient distributed optimizers such as DiLoCo reduce synchronization costs by letting workers perform many local updates before aggregat

local-aiarxiv-cs-lg
28 May 2026
Model Releases

Parameter-Efficient Generative Modeling with Controlled Vector Fields

DGX agent

arXiv:2605.28267v1 Announce Type: new Abstract: We introduce a continuous-time generative modeling framework, motivated by the Chow-Rashevskii theorem, that builds expressive flows from a small set of

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Reproducibility Below the Rerun-Stability Baseline

DGX agent

arXiv:2605.27440v1 Announce Type: cross Abstract: Small changes to how a buyer phrases a question -- 'best CRM' vs 'top CRM' vs 'best CRM for a SaaS startup' -- produce substantially different brand r

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Particle-Guided Diffusion Models for Partial Differential Equations

DGX agent

arXiv:2601.23262v2 Announce Type: replace Abstract: We introduce a guided stochastic sampling method that augments sampling from diffusion models with physics-based guidance derived from partial diffe

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

PAST2HARM: A Simple Adaptive Past Tense Attack for Jailbreaking Multimodal AI

DGX agent

arXiv:2605.27545v1 Announce Type: new Abstract: Jailbreak attacks on multimodal AI systems remain underexplored, even though unsafe image generation can have more severe consequences than unsafe text

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Patched-DeltaNet: Token-Level Event-Driven Memory for Linear-Time Anomaly Detection

DGX agent

arXiv:2605.27992v1 Announce Type: new Abstract: Time series anomaly detection is critical for maintaining the reliability of mission-critical systems. While Transformer-based models like PatchTST have

model-releasesarxiv-cs-lg
28 May 2026
Research

Path Channels and Plan Extension Kernels: a Mechanistic Description of Planning in a Sokoban RNN

DGX agent

arXiv:2506.10138v3 Announce Type: replace-cross Abstract: We partially reverse-engineer a convolutional recurrent neural network (RNN) trained with model-free reinforcement learning to play the box-pu

researcharxiv-cs-ai
28 May 2026
Research

Pattern Recognition Tasks with Personalized Federated Learning

DGX agent

arXiv:2605.27816v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) constitutes a novel paradigm that tailors Machine Learning (ML) models to individual clients, thereby furnishing p

researcharxiv-cs-cv
28 May 2026
Model Releases

PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience in Minecraft

DGX agent

arXiv:2605.27762v1 Announce Type: new Abstract: We present PEAM, a Parametric Embodied Agent Memory framework in Minecraft that transforms agent memory from inference-time retrieval into parameter-res

model-releasesarxiv-cs-ai
28 May 2026
Research

PEAR: Equal Area Weather Forecasting on the Sphere

DGX agent

arXiv:2505.17720v3 Announce Type: replace Abstract: Artificial intelligence is rapidly reshaping the natural sciences, with weather forecasting emerging as a flagship AI4Science application where mach

researcharxiv-cs-lg
28 May 2026
Model Releases

PEAR: Pairwise Evaluation for Automatic Relative Scoring in Machine Translation

DGX agent

arXiv:2601.18006v2 Announce Type: replace Abstract: We present PEAR (Pairwise Evaluation for Automatic Relative Scoring), a supervised quality estimation (QE) metric family that reframes reference-fre

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective

DGX agent

arXiv:2605.28819v1 Announce Type: cross Abstract: Parameter-efficient finetuning (PEFT) has become the standard approach for adapting large language models, yet evaluations largely emphasize downstrea

model-releasesarxiv-cs-cl
28 May 2026
Applications

Performance and Explainability Requirements of Evolutionary Algorithms in Real-World Physics-Informed Optimization

DGX agent

arXiv:2605.28164v1 Announce Type: cross Abstract: Evolutionary computation offers a variety of tools to solve complex real-world optimization problems. However, research often focuses on smaller, simp

applicationsarxiv-cs-ai
28 May 2026
Model Releases

Periodic RoPE for Infinite Context LLMs

DGX agent

arXiv:2605.27980v1 Announce Type: cross Abstract: The ability to process ultra-long contexts is crucial for large language models (LLMs) to perform long-horizon tasks. While recent efforts have extend

model-releasesarxiv-cs-ai
28 May 2026
Tools

Perplexity Computer is now available inside Microsoft Excel, Word, PowerPoint, and Outlook. Orchestrate across work with Computer directly i…

DGX agent

Perplexity Computer is now available inside Microsoft Excel, Word, PowerPoint, and Outlook. Orchestrate across work with Computer directly in the side panel of your app to draft documents, model, buil

toolsperplexity--x
28 May 2026
Model Releases

Personal Visual Memory from Explicit and Implicit Evidence

DGX agent

arXiv:2605.28806v1 Announce Type: cross Abstract: Long-term memory is increasingly important for personalized AI agents, yet existing benchmarks and methods remain largely text-centric. Even when imag

model-releasesarxiv-cs-cl
28 May 2026
Agents

Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis

DGX agent

arXiv:2605.28037v1 Announce Type: new Abstract: Prompt-based personality control is a key technique for designing large language model (LLM) dialogue agents that behave consistently across social cont

agentsarxiv-cs-cl
28 May 2026
Model Releases

Personalized Observation Normalization for Federated Reinforcement Learning in Simulation Environments with Heterogeneity

DGX agent

arXiv:2605.27385v1 Announce Type: cross Abstract: Federated reinforcement learning (FedRL) enables multiple agents to collaboratively train a global policy without sharing raw data, making it ideal fo

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models

DGX agent

arXiv:2503.01829v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate persuasive capabilities that rival human-level persuasion. While these capabilities can be used for s

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

PetroBench: A Benchmark for Large Language Models in Petroleum Engineering

DGX agent

arXiv:2605.28032v1 Announce Type: new Abstract: Large Language Models are increasingly applied in the petroleum industry, highlighting the need for a domain-specific evaluation framework. This study d

model-releasesarxiv-cs-ai
28 May 2026
Research

PhAME: Phenotype-Aware Molecular Editing via Latent Diffusion

DGX agent

arXiv:2605.28226v1 Announce Type: new Abstract: Small-molecule drug discovery requires simultaneous optimization of numerous properties of candidate molecules. These properties can be investigated thr

researcharxiv-cs-lg
28 May 2026
Safety

Picid: A Modular Evaluation Infrastructure for Reproducible PHM Across Tasks and Domains

DGX agent

arXiv:2605.28345v1 Announce Type: new Abstract: Progress in Prognostics and Health Management (PHM) is hindered by the lack of standardized and reusable evaluation practices across tasks, datasets, an

safetyarxiv-cs-ai
28 May 2026
Model Releases

PINE: Pruning Boosted Tree Ensembles with Conformal In-Distribution Prediction Equivalence

DGX agent

arXiv:2605.28068v1 Announce Type: new Abstract: Tree ensembles are machine learning models with strong predictive performance and interpretability, and remain widely used for tabular data. Standard pr

model-releasesarxiv-cs-lg
28 May 2026
Agents

PIRS: Physics-Informed Reward Shaping for SAC-Based Building Energy Management

DGX agent

arXiv:2605.28232v1 Announce Type: new Abstract: Occupant comfort and grid-aware energy efficiency are competing objectives whose joint optimization depends critically on how reward functions are speci

agentsarxiv-cs-ai
28 May 2026
Model Releases

Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona (Rashi Shrivastava/Forbes)

DGX agent

Rashi Shrivastava / Forbes: Pittsburgh-based Gray Swan, which stress-tests AI models for top frontier AI labs, raised a 40M Series A at a 200M valuation co-led by Wing VC and Madrona — Gray Swan works

model-releasestechmeme
28 May 2026
Agents

Plan Before Search: Search Agents Need Plan

DGX agent

arXiv:2605.28354v1 Announce Type: new Abstract: Training large language models as retrieval-augmented reasoning agents typically combines reinforcement learning with an SFT cold start distilled from a

agentsarxiv-cs-ai
28 May 2026
Research

Planning a Community Approach to Diabetes Care in Low- and Middle-Income Countries Using Optimization

DGX agent

arXiv:2305.06426v2 Announce Type: replace Abstract: Diabetes is a global health priority, especially in low- and-middle-income countries, where over 50% of premature deaths are attributed to high bloo

researcharxiv-cs-ai
28 May 2026
Model Releases

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

DGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

model-releasesarxiv-cs-ai
28 May 2026
Research

Playing with Words, Improving with Rewards: Training Language Models for Creative Association

DGX agent

arXiv:2605.27832v1 Announce Type: new Abstract: Large Language Models (LLMs) are being applied to increasingly difficult problems and use cases. To navigate their vast solution spaces effectively, LLM

researcharxiv-cs-cl
28 May 2026
Research

PLS in the Mirror of Self-Attention

DGX agent

arXiv:2605.28592v1 Announce Type: new Abstract: This note provides an interesting observation on casting partial least square (PLS) as a linearized self-attention so that PLS may be studied within the

researcharxiv-cs-lg
28 May 2026
Model Releases

Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

DGX agent

arXiv:2601.15015v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown promising results in active flow control (AFC), yet progress in the field remains difficult to assess as exist

model-releasesarxiv-cs-lg
28 May 2026
Local Ai

PocketGS: On-Device Training of 3D Gaussian Splatting for High Perceptual Modeling

DGX agent

arXiv:2601.17354v5 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) enables real-time rendering, its training demands workstation-level compute and memory, making mobile deployment

local-aiarxiv-cs-cv
28 May 2026
Model Releases

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation

DGX agent

arXiv:2605.28237v1 Announce Type: cross Abstract: Real-world navigation is fundamentally driven by Points of Interest (POIs), yet reaching a precise POI remains a critical 'final-meters' challenge. Ex

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

PointQ-Bench: Benchmarking Diagnostic and Interpretable Point Cloud Quality Assessment

DGX agent

arXiv:2605.28241v1 Announce Type: new Abstract: Point cloud quality plays a critical role in 3D acquisition, reconstruction, rendering, and perception, yet existing point cloud quality assessment (PCQ

model-releasesarxiv-cs-cv
28 May 2026
Research

Poison with Style: A Practical Poisoning Attack on Code Large Language Models

DGX agent

arXiv:2605.27631v1 Announce Type: cross Abstract: Code Large Language Models (CLLMs) serve as the core of modern code agents, enabling developers to automate complex software development tasks. In thi

researcharxiv-cs-lg
28 May 2026
Model Releases

PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Management

DGX agent

arXiv:2605.27887v1 Announce Type: new Abstract: LLMs have shown strong performance across diverse financial tasks, yet portfolio management (PM), a critical financial decision-making task, remains poo

model-releasesarxiv-cs-ai
28 May 2026
Safety

Position: Retire the 'Positive Backdoor' Label -- Secret Alignment Requires Strict and Systematic Evaluation

DGX agent

arXiv:2605.28597v1 Announce Type: cross Abstract: This position paper argues that the AI/ML community should stop overclaiming and retire the label 'positive backdoor,' and instead treat trigger-activ

safetyarxiv-cs-ai
28 May 2026
Tools

Power users account for a large share of AI activity, and the gap is widening.

DGX agent

Power users represent a disproportionately large portion of activity in AI platforms, with their share of total usage growing over time. This trend indicates increasing concentration of AI tool usage

toolscursor--x
28 May 2026
Research

Preference-Shaped Expected Hypervolume and R2 Improvement: Exact Computation and Monotonicity

DGX agent

arXiv:2605.28746v1 Announce Type: cross Abstract: This paper studies preference-shaped expected improvement criteria for Bayesian multiobjective optimization. We consider two indicator families which

researcharxiv-cs-ai
28 May 2026
Research

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking

DGX agent

arXiv:2605.27712v1 Announce Type: new Abstract: Long reasoning traces need reliability estimates before final answers are known. We study prefix-conditioned eventual-success estimation, P(y=1 mid o_{1

researcharxiv-cs-ai
28 May 2026
Agents

President and Head of AI at Replit, @pirroh is the architect of Replit Agent and former Head of Applied Research at Google X. See him take t…

DGX agent

President and Head of AI at Replit, @pirroh is the architect of Replit Agent and former Head of Applied Research at Google X. See him take the stage with @refikanadol on day two of Vibecon. NYC, June

agentsreplit--x
28 May 2026
Model Releases

Pressure-Testing Deception Probes in LLMs: Scaling, Robustness, and the Geometry of Deceptive Representations

DGX agent

arXiv:2605.27958v1 Announce Type: cross Abstract: Linear probes trained on LLM activations are increasingly proposed as deception-detection metrics, yet report AUROC exceeding 0.96 on clean benchmarks

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation

DGX agent

arXiv:2605.28634v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising paradigm for generalist robotic policies, yet their adaptation is hindered by data inefficiency an

tutorialsarxiv-cs-ro
28 May 2026
Applications

Principled Algorithms for Optimizing Generalized Metrics in Multi-Label Learning

DGX agent

arXiv:2605.28767v1 Announce Type: new Abstract: Many real-world classification tasks require predicting multiple labels per instance, necessitating the optimization of complex evaluation metrics such

applicationsarxiv-cs-lg
28 May 2026
Model Releases

PrionNER: A Named Entity Recognition Dataset for Prion Disease Biomedical Literature

DGX agent

arXiv:2605.28375v1 Announce Type: new Abstract: Prion diseases are rare, rapidly progressive, and fatal neurodegenerative disorders that remain difficult to diagnose, particularly in their early stage

model-releasesarxiv-cs-cl
28 May 2026
← Previous
1…10291030103110321033…1898
Next →