Sightings
/elsewhere/sightings/ I have a new camera (a Canon R6 Mark II) so I'm taking a lot more photos of birds. I share my best wildlife photos on iNaturalist, and based on yesterday's successful prototype I
Knowledge catalogue
/elsewhere/sightings/ I have a new camera (a Canon R6 Mark II) so I'm taking a lot more photos of birds. I share my best wildlife photos on iNaturalist, and based on yesterday's successful prototype I
This is actually a version of an alignment problem. Humans have background beliefs (don’t waste large sums of money without telling me) and Claude doesn’t respect those. Caveat emptor. THIS GUY ACCIDE
We are honored to be featured in the latest @TwoMinutePapers video! You all can watch the full video here: https://youtu.be/QzZ4VwDHAT4 Here’s a short clip from it: Media What happens when you put com
arXiv:2604.28001v1 Announce Type: new Abstract: Integrating multimodal foundation models into enterprise ecosystems presents a fundamental software architecture challenge. Architects must balance comp
arXiv:2604.28173v1 Announce Type: new Abstract: Effective human behavior modeling requires a representation of the human body movement that capitalizes on its compositionality. We propose a hierarchic
arXiv:2604.27434v1 Announce Type: cross Abstract: Federated learning (FL) is a popular distributed learning paradigm in machine learning, which enables multiple clients to collaboratively train models
On April 7, 2026, Anthropic did something unprecedented in the history of artificial intelligence: The company announced that it had built its most capable model ever and would not be releasing it to
Lauren Forristal / TechCrunch: Amazon debuts “Join the chat”, an AI-powered feature that lets users ask questions about products and get conversational audio responses generated in real time — Amazon
arXiv:2604.27644v1 Announce Type: cross Abstract: We propose a paradigm shift from learning to answer to learning to question: can a language model generate verifiable problems, solve them, and turn t
arXiv:2604.27543v1 Announce Type: new Abstract: Evaluating English ASR systems for conversational AI applications remains difficult, as many publicly available corpora are either pre-segmented into sh
arXiv:2604.27105v1 Announce Type: new Abstract: Analyzing mutual gaze (MG) and joint attention (JA) is critical in developmental psychology but traditionally relies on labor-intensive manual coding. A
arXiv:2604.27394v1 Announce Type: cross Abstract: Conditional Average Treatment Effect (CATE) estimation in practice demands three properties simultaneously: heterogeneous effects au(x), calibrated un
This post appears to be a sports-related social media update about an upcoming Hawks game, though the message is incomplete and cuts off mid-sentence before providing specific details about the score
OpenAI announced a feature that allows users to quickly import their existing workflow configurations, including settings, plugins, agents, and project configurations, into Codex with minimal effort.
arXiv:2604.27591v1 Announce Type: cross Abstract: Video moment retrieval is the task of retrieving specific segments of a video corresponding to a given text query. Recent studies have been conducted
Code with Claude, our developer conference, returns next week. Whether you're just getting started with Claude Code or you've been building for a while, there's a session for you. Register for the liv
Yohei Nakajima announced the launch of Cofounder 2 on May 4th via X. Cofounder is an AI agent tool designed to assist with business and startup tasks. The announcement was shared on social media to in
arXiv:2508.13316v2 Announce Type: replace Abstract: We consider the problem of designing constraint-aware flow matching (FM) models that address the issue of constraint violations commonly observed in
arXiv:2604.27707v1 Announce Type: new Abstract: Current agentic memory systems (vector stores, retrieval-augmented generation, scratchpads, and context-window management) do not implement memory: they
arXiv:2604.27883v1 Announce Type: cross Abstract: In modern parametric model training, full-batch gradient descent (and its variants) suffers due to progressively stronger biasing towards the exact re
arXiv:2603.09117v2 Announce Type: replace-cross Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) significantly enhances large language models (LLMs) reasoning but severely suffers from
arXiv:2604.27004v1 Announce Type: cross Abstract: We propose EdgeSpike, a co-designed spiking neural network (SNN) framework for autonomous low-power sensing in edge Internet of Things (IoT) architect
arXiv:2604.27275v1 Announce Type: cross Abstract: Large language model (LLM) reading assistants are increasingly used in settings that require interpretation rather than simple retrieval. In these con
arXiv:2604.27195v1 Announce Type: new Abstract: Accurate prediction of conversion from Mild Cognitive Impairment (MCI) to Alzheimers Diseases (AD) is essential for early intervention, however, develop
Event with @googlegemma next week! Hang out with the fellow builders and members of the Gemma + LM teams at LM Studio HQ in NYC. When: Monday, May 11th RSVP: required, link below High likelihood of pi
arXiv:2604.27590v1 Announce Type: new Abstract: Recent advances in 3D reconstruction and neural rendering,particularly 3D Gaussian Splatting, make it feasible and simple to edit 3D scenes and re-rende
arXiv:2604.28102v1 Announce Type: new Abstract: Solving practical multi-depot vehicle routing problems (MDVRP) is a challenging optimization task central to modern logistics, increasingly driven by e-
arXiv:2604.27353v1 Announce Type: new Abstract: Gait recognition has emerged as a compelling biometric modality for surveillance and security applications, offering inherent advantages such as non-int
arXiv:2604.28193v1 Announce Type: new Abstract: Reconstructing 3D scenes from sparse, unposed images remains challenging under real-world conditions with varying illumination and transient occlusions.
arXiv:2604.27918v1 Announce Type: new Abstract: Existing talking avatar methods typically adopt an image-to-video pipeline conditioned on a static reference image within the same scene as the target g
@GoogleAIStudio And this sample project was created on Canvas in @GeminiApp. It’s a high-speed rhythm game where you tap to the beat and collect power-ups to remix the track. Watch as the numbers appe
@GoogleAIStudio @GeminiApp We can’t wait to see where your creativity takes you. Vibe code your countdown idea in @GoogleAIStudio or Canvas in @GeminiApp, then submit it here: http://goo.gle/codetheco
arXiv:2509.16248v3 Announce Type: replace-cross Abstract: This paper presents GRAPHMEND, a high-level compiler technique that eliminates FX graph breaks in PyTorch 2 programs. Although PyTorch 2 intro
Hi Singapore 🇸🇬, meet Codex 🩵 With Codex, ANYONE can build and create. We’re turning that energy up this May. We’re a diamond sponsor 💎 at @aiDotEngineer Singapore, at a bunch of events, and hosting s
arXiv:2604.27903v1 Announce Type: new Abstract: The rapid evolution of generative models has enabled the creation of highly realistic and diverse synthetic images, posing significant challenges to rel
arXiv:2604.27790v1 Announce Type: cross Abstract: Generative AI is being increasingly integrated into web search for the convenience it provides users. In this work, we aim to understand how generativ
This Reddit post likely discusses the technical architecture of Google's Imagen 2 text-to-image model. Imagen generates images in pixel space , contrasting with latent diffusion approaches. The discus
arXiv:2501.19143v2 Announce Type: replace Abstract: As the cornerstone of artificial intelligence, machine perception confronts a fundamental threat posed by adversarial illusions. These adversarial a
arXiv:2604.27540v1 Announce Type: new Abstract: Scientific reasoning rarely stops at what is directly observable; it often requires uncovering hidden structure from data. From estimating reaction cons
arXiv:2604.27891v1 Announce Type: new Abstract: Agent orchestration frameworks -- LangGraph, CrewAI, Google ADK, OpenAI Agents SDK, and others -- place an external orchestrator above the LLM, tracking
Tool: iNaturalist Sightings I wanted to see my iNaturalist observations - across two separate accounts - grouped by when they occurred. I'm camping this weekend so I built this entirely on my phone us
This post suggests that OpenAI has achieved recursive self-improvement capabilities in Codex, their code generation model, potentially enabling the system to iteratively enhance its own performance. T
arXiv:2504.14602v2 Announce Type: replace-cross Abstract: The natural interaction and control performance of lower limb rehabilitation robots are closely linked to biomechanical information from vario
arXiv:2604.28010v1 Announce Type: cross Abstract: We reframe clinician overrides of clinical AI recommendations as implicit preference data - the same signal structure exploited by reinforcement learn
arXiv:2604.27295v1 Announce Type: new Abstract: Learning rate scheduling has evolved from the single global fixed rate of early SGD to sophisticated layer-wise adaptive strategies. We systematize this
arXiv:2604.27063v1 Announce Type: new Abstract: Continual learning agents with finite capacity must balance acquiring new knowledge with retaining the old. This requires controlled forgetting of knowl
arXiv:2604.26973v1 Announce Type: cross Abstract: Multiobjective optimization remains challenging for many scientific and engineering problems due to the need to balance convergence, diversity, and co
arXiv:2604.27624v1 Announce Type: cross Abstract: Large Language Models (LLMs) can strongly shape social discourse, yet datasets investigating how LLM outputs vary across controlled social and context
arXiv:2604.26984v1 Announce Type: new Abstract: Representational collapse, where embeddings become anisotropic and lose multi-scale structure, can erode downstream performance long before performance
New paper (on an old AI) tests o1 against doctors on medical benchmarks & real ER cases: “across a variety of scenarios and applications, the large language model outperformed both human physicians an
arXiv:2604.27031v1 Announce Type: cross Abstract: In a continual learning setting, we require a model to be plastic enough to learn a new task and stable enough to not disturb previously learned capab
On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session! Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to
Once again IP thieves accusing other IP thieves of IP thievery! The nerve! @ChrisRMcGuire Anthropic claimed OpenAI violated their terms of service by using their API (I suspect they would call it “dis
arXiv:2604.27599v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for recommendation reranking, but their listwise predictions can depend on the order in which candi
Our CEO @jerryjliu0 in @VentureBeat , on what's actually changing in the LLM stack: 'We've really identified that there's a core set of data that has been locked up in all these file format containers
arXiv:2604.27870v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) are widely assumed to be translation-invariant, yet standard architectures exhibit a startling fragility: even a si
arXiv:2604.26991v1 Announce Type: cross Abstract: Recent advances in data-centric medical AI have produced highly accurate diagnostic systems, but the emphasis on data curation and performance metrics
arXiv:2604.27182v1 Announce Type: cross Abstract: Time-series data augmentation plays a crucial role in regression-oriented forecasting tasks, where limited data restricts the performance of deep lear
arXiv:2604.28055v1 Announce Type: cross Abstract: Individualized Alzheimer's disease (AD) progression prediction requires models that use irregular visits, account for censoring, avoid diagnostic leak
arXiv:2604.27472v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models advance robotic control via strong visual-linguistic priors. However, existing VLAs predominantly frame pretraining