General Hazard Detection
arXiv:2605.23304v1 Announce Type: new Abstract: Hazard, as an abstract concept, is typically defined through cognitive-level logical reasoning rather than concrete examples. In contrast, existing haza
Knowledge catalogue
arXiv:2605.23304v1 Announce Type: new Abstract: Hazard, as an abstract concept, is typically defined through cognitive-level logical reasoning rather than concrete examples. In contrast, existing haza
arXiv:2605.23159v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) is expected to transform work, but less is known about how firms reorganize labor demand as the technology dif
arXiv:2605.23555v1 Announce Type: new Abstract: This paper addresses the challenge of reconstructing photorealistic and animatable 3D human avatars from monocular videos. While existing methods rely o
arXiv:2605.23888v1 Announce Type: new Abstract: We introduce a new approach to high-fidelity 3D scene reconstruction from multi-view RGB images that tightly couples reconstruction with a strong genera
arXiv:2605.23238v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as economic agents in marketplaces, auctions, and bidding settings. Anticipating their behavior i
arXiv:2605.23903v1 Announce Type: new Abstract: Camera-controlled video generation has achieved remarkable progress in recent years. However, existing video-to-video re-rendering methods primarily rel
arXiv:2508.14083v3 Announce Type: replace-cross Abstract: The ubiquity of missing data in urban intelligence systems, attributable to adverse environmental conditions and equipment failures, poses a s
arXiv:2605.23327v1 Announce Type: new Abstract: Lane detection stands as a crucial perception task in autonomous driving and advanced driver assistance systems. However, existing methods still degrade
arXiv:2510.04567v2 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) are powerful tools for processing relational data but often struggle to generalize to unseen graphs, giving rise
Given how much of the original 'bottle of water per generated email' water estimate came from guesses at the architecture of GPT-4, it would be very much in @OpenAI's interest to publish the architect
Globally renowned digital artist behind the iconic Windows 10 wallpaper seen by billions, @gmunk’s psychedelic, atmospheric work spans installations, music videos, title sequences, and immersive film.
arXiv:2605.23602v1 Announce Type: new Abstract: Existing 3DGS methods effectively render high-quality novel views in clear-day scenes. However, they struggle with night scenes, particularly in glow re
arXiv:2504.09846v2 Announce Type: replace-cross Abstract: Frequent and long-term exposure to hyperglycemia increases the risk of chronic complications, including neuropathy, nephropathy, and cardiovas
arXiv:2605.23183v1 Announce Type: cross Abstract: Contemporary glioma diagnosis integrates molecular features with histopathology to guide clinical decision-making. However, in clinical settings, dive
arXiv:2605.23551v1 Announce Type: cross Abstract: A goal-conditioned reinforcement learning agent exploring an environment will see a wealth of information throughout a trajectory, most of which is di
/goal is really insane! It's how you can get the most out of coding agents today. For efficiency, I find it works best when you do planning before /goal. This ensures the agent has the right context a
arXiv:2605.23892v1 Announce Type: cross Abstract: Visual geometry transformers have become powerful architectures for multi-view 3D reconstruction, enabling joint prediction of multiple 3D attributes
🚨 Google DeepMind CEO Sir Demis Hassabis: “Today’s systems, are nowhere near [AGI]. Doesn’t matter how many Erdős problems you solve… I think it’s far, far from what a true invention or someone like a
arXiv:2602.11629v2 Announce Type: replace Abstract: Graph Prompt Learning (GPL) has recently emerged as a promising paradigm for downstream adaptation of pre-trained graph models, mitigating the misal
arXiv:2602.00979v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as educational agents for automatic short answer grading (ASAG) in real-world education
arXiv:2605.22963v1 Announce Type: cross Abstract: Large Language Models (LLMs) are optimized to produce distributionally plausible continuations rather than to explicitly verify whether generated prop
arXiv:2605.23696v1 Announce Type: new Abstract: Effectively managing Air Traffic Control Officer (ATCO) workload is crucial in maintaining operational safety. Group supervisors use tools that estimate
arXiv:2508.10651v3 Announce Type: replace Abstract: We present a novel approach for graph classification based on tabularizing graph data via new variants of the Weisfeiler-Leman algorithm and then ap
Great that so many of you have enjoyed my deconstructions of Elon, Zuckerberg, and that OpenAI dude, but this is much more important: ⚠️⚠️⚠️i don’t think most people understand the implications of the
Grok Build improving every day, 7 days a week Bug fixes shipping to Grok Build 0.1.219 (release notes will be available in the TUI) - fixing usage limit bugs with prompt caching - fix layout-shifted c
Grok Build is now available in Beta for all SuperGrok and X Premium+ users. Use Plan Mode, create images and videos with Imagine, and build automations or orchestrators with the CLI. Visit http://x.ai
Grok Build is still in beta for another month or so, but is already quite useful for production tasks Try it out! Favorite features: - <1 second web/X search - Editing and creating assets with Imagine
Grok foundation model V9-Medium (1.5T) has finished training. Evals look good. A lot of Cursor data was added in supplementary training and there is more to come. Fine-tuning is underway and reinforce
Groove Jones is a customer or partner of ComfyUI, as featured on the ComfyUI customers page. The post likely highlights Groove Jones's use of ComfyUI, a node-based UI for Stable Diffusion and other ge
. @GrooveJonesXR needed to deliver the impossible: giant NFL-licensed Crocs parachuting into Dick's Sporting Goods parking lots; hyper-realistic, multi-location, vertical 9:16 on a holiday deadline. A
arXiv:2602.12316v2 Announce Type: replace Abstract: Frontier AI systems are increasingly capable and deployed in high-stakes multi-agent environments. However, existing AI safety benchmarks largely ev
arXiv:2602.05202v2 Announce Type: replace Abstract: Aligning video generative models with human preferences remains challenging: current approaches rely on Vision-Language Models (VLMs) for reward mod
Ha. @demishassabis and @DarioAmodei, who are running the two best AI labs in the world both have PhDs. Jim Simons, who founded one of the biggest hedge funds in the world had a PhD. Eric Schmidt, Goog
arXiv:2605.23572v1 Announce Type: cross Abstract: In the competitive landscape of sponsored search, balancing retrieval quality with production latency is a critical challenge. While large retrieval m
This article defines and clarifies key terminology related to AI agents, including the concepts of 'harness' and 'scaffold,' which are important architectural and operational components in building an
arXiv:2605.23043v1 Announce Type: new Abstract: Agentic text-simulation systems write in sequence, with each item becoming possible context for later steps. That makes uncertainty path-dependent: an e
Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official/autonomous-ai-agents/openhands` Reminder: You can already d
arXiv:2605.23190v1 Announce Type: new Abstract: Machine-generated texts (MGTs) produced by large language models (LLMs) are increasingly prevalent across various applications, while their potential mi
arXiv:2605.23821v1 Announce Type: new Abstract: We propose a distributional theory of how hypernymy -- the ``is-a'' relation between general and specific concepts -- is encoded geometrically in langua
arXiv:2605.23422v1 Announce Type: new Abstract: Learning high-quality oblique decision trees remains a significant challenge due to the discrete and non-convex nature of split optimization. We present
arXiv:2605.23889v1 Announce Type: new Abstract: Online 3D reconstruction requires estimating camera pose and scene geometry under strict causal and bounded-memory constraints. Existing methods often s
Gary Marcus humorously suggests changing one's social media name to 'Kekius Maximus,' likely commenting on internet culture, meme terminology, or the absurdity of online personas. The post appears to
arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio
arXiv:2605.22880v1 Announce Type: cross Abstract: As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence cam
arXiv:2605.23628v1 Announce Type: new Abstract: Multi-task benchmarks have become a central pillar of machine learning research, yet their growing influence has incentivised benchmark gaming -- strate
arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho
Check Point Research: How Iranian threat actor Nimbus Manticore used techniques like AI-assisted malware development and SEO poisoning to target companies during the US-Iran war — Key Findings — The I
arXiv:2605.23583v1 Announce Type: cross Abstract: Inverse Kinematics (IK) plays a critical role in robotic motion planning and control. The IK solutions of a robot manipulator could be done by convent
arXiv:2603.10067v2 Announce Type: replace-cross Abstract: Muon has recently shown promising results in LLM training. In this work, we study how to further improve Muon. We argue that Muon's orthogonal
The post reviews the Hugging Face Reachy Mini robot, praising its build quality, IDE, and developer ecosystem for desktop companion robotics applications. It highlights the robot's support for agentic
arXiv:2605.22940v1 Announce Type: cross Abstract: Deep learning is increasingly viewed as a dynamical process in parameter space, yet many existing theories still treat training as a closed optimizati
arXiv:2605.23867v1 Announce Type: cross Abstract: Large language models (LLMs) have the potential to aid and improve human decision-making in classification tasks, not only by providing fairly accurat
arXiv:2605.23320v1 Announce Type: new Abstract: Ventilator decision support requires sequential decisions that track evolving physiology and disease trajectories while respecting safety boundaries and
arXiv:2605.23403v1 Announce Type: new Abstract: Statistical downscaling is a crucial component of the weather modeling field, where high-resolution outputs must be reconstructed from coarse-resolution
Simon Willison reflects on the disconcerting experience of unknowingly reading AI-written emails presented as human-authored, comparing it to deliberate deception. The post raises concerns about authe
An Ars Technica article examining the cultural pressure to complete highly-acclaimed games like The Witcher 3, arguing that a game being objectively good doesn't obligate players to enjoy it or finish
Gary Marcus comments on how widespread reporting of similar issues or problems across multiple companies can lead to a collapse of confidence or a 'bubble pop' in a particular sector or narrative. The
If I wrote a short, cheap, electronic sequel to Taming Silicon Valley, perhaps with a title like “Getting to an AI that is good for all of humanity” (perhaps donating proceeds) would you read it and g
If you live in the UK, and care about the risk from superintelligence, please read below! URGENT: The UK Parliament just selected 20 MPs to introduce a bill of their choosing. This is an unprecedented
(I'm firmly on team red/green TDD for agent code, I like having a test suite that protects against them breaking old features when they make new changes - https://simonwillison.net/guides/agentic-engi