Temporal Graph Pattern Machine
arXiv:2601.22454v3 Announce Type: replace Abstract: Temporal graph learning is pivotal for deciphering dynamic systems, where the core challenge lies in explicitly modeling the underlying evolving pat
Knowledge catalogue
arXiv:2601.22454v3 Announce Type: replace Abstract: Temporal graph learning is pivotal for deciphering dynamic systems, where the core challenge lies in explicitly modeling the underlying evolving pat
arXiv:2606.23120v1 Announce Type: new Abstract: The goal of source-free domain adaptation (SFDA) for time-series data is to transfer knowledge from a pre-trained source model to an unlabeled target do
arXiv:2606.23212v1 Announce Type: new Abstract: Despite modeling temporal motion, dynamic 3D Gaussian Splatting (3DGS) methods still inherit a static densification strategy that is ill-suited for dyna
arXiv:2505.17782v4 Announce Type: replace Abstract: Monitoring volcanic activity is of paramount importance to safeguarding lives, infrastructure, and ecosystems. However, only a small fraction of kno
That's it. Point it at a channel, give it a task, and let it work. It's in beta on Slack today for Claude Enterprise and Team customers. More surfaces coming soon! Learn more: https://www.anthropic.co
arXiv:2606.22686v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) rely on extensive safety alignment, yet the mechanistic basis of refusal remains opaque. In this work, we investig
arXiv:2603.01250v3 Announce Type: replace Abstract: Breast cancer is the most frequently diagnosed malignancy among women worldwide and a leading cause of cancer-related mortality. Dynamic contrast-en
arXiv:2606.21008v1 Announce Type: cross Abstract: The metanym game is a competitive word game for LLMs that measures structural intelligence against established cognitive-science constructs. No conten
arXiv:2606.22574v1 Announce Type: new Abstract: While synthetic data generation resolves the manual labeling bottleneck in computer vision, minimizing the syn-to-real domain gap requires optimizing re
arXiv:2606.21611v1 Announce Type: new Abstract: Mathematical search problems present a unique challenge for Reinforcement Learning (RL) due to vast search spaces and sparse rewards. In previous works,
This is a new paradigm for interacting with Claude that is significantly more 'inline' with all the other human activity org-wide. Once you do all of the under the hood engineering work to make this '
Claude Everywhere represents Anthropic's expansion of Claude's capabilities across multiple platforms and interfaces, leveraging Claude Code as its underlying technology to maintain consistent code-wr
arXiv:2606.22969v1 Announce Type: new Abstract: Predicting the behavior of dynamical systems (DS) beyond the dynamical and parameter regimes observed in training is a pivotal and essentially unresolve
arXiv:2606.22029v1 Announce Type: new Abstract: Fingerprints are the most widely deployed biometric. Verifying whether two impressions come from the same finger typically relies on minutiae, small lan
arXiv:2606.22370v1 Announce Type: new Abstract: Recent advances in video generation have made minute-level synthesis possible; however, generating long videos remains challenging due to error accumula
arXiv:2603.25260v2 Announce Type: replace Abstract: LiDAR point clouds are fundamental to various applications, yet the extreme sparsity of high-precision geometric details hinders efficient context m
arXiv:2606.22782v1 Announce Type: new Abstract: The proliferation of IoT devices has fueled distributed edge systems to collect vast amounts of sensitive data, creating fertile ground for on-device ma
arXiv:2606.22589v1 Announce Type: new Abstract: Ever since the advent of foundation models and the pre-training-finetuning paradigm, there have been numerous efforts to merge multiple task-specific ex
arXiv:2606.20852v1 Announce Type: new Abstract: Inference-time engineering can alter model behavior without fine-tuning. However, its utility for improving diagnostic performance in medical vision-lan
arXiv:2603.06608v2 Announce Type: replace-cross Abstract: The research community lacks a middle ground between StarCraft II full game and its mini-games. The full-game's sprawling state-action space r
arXiv:2606.19328v2 Announce Type: replace Abstract: Preference-based RL provides an approach to learning reward models from pairwise comparisons of behaviors, bypassing the need for explicit reward de
arXiv:2606.21223v1 Announce Type: new Abstract: Reliable localization is essential for intelligent transportation systems (ITS), including autonomous vehicles, quadruped last-mile carriers, and infras
arXiv:2606.22976v1 Announce Type: new Abstract: In this paper, we propose using random walks on graphs as a verifiable sandbox to study different parallel sampling strategies in masked diffusion model
arXiv:2606.21847v1 Announce Type: new Abstract: Low-rank decomposition serves as a promising compression paradigm for large language models, however, rank allocation remains challenging: manual rules
arXiv:2606.21661v1 Announce Type: new Abstract: Generating a coherent multi-shot video requires structured cross-shot memory. Subject appearance, scene context, and speaker identity must persist acros
arXiv:2606.23050v1 Announce Type: new Abstract: Recently, end-to-end OCR models, exemplified by DeepSeek OCR, have once again thrust OCR into the spotlight. A widely held view is that employing a larg
arXiv:2606.23610v1 Announce Type: new Abstract: Video diffusion models have enabled remarkable progress in video generation and editing. However, content preservation remains a core challenge: existin
arXiv:2606.23543v1 Announce Type: cross Abstract: Scaling reinforcement learning for visual mathematical reasoning requires more than generating harder questions: as data volume grows, the reward labe
Protecting sensitive data used with AI is a critical part of our commitment to providing advanced and secure cloud infrastructure. Confidential Computing cryptographically protects data in use in hard
arXiv:2606.23327v1 Announce Type: new Abstract: Video editing has become essential in digital media creation, yet existing automated systems are restricted to short segment processing and domain-speci
arXiv:2606.17710v2 Announce Type: replace Abstract: Medical vision-language models report strong chest radiograph accuracy, and this is increasingly read as evidence that they use the image. That infe
arXiv:2606.23062v1 Announce Type: cross Abstract: We introduce VolHuMe, a dataset of high-quality 4D human scans captured with a state-of-the-art volumetric studio using 64 RGB and 32 depth cameras. V
arXiv:2606.23145v1 Announce Type: new Abstract: Operational event-detection systems are rarely assessed by pointwise accuracy alone. In anomaly detection, changepoint detection, and warning systems, t
We're launching Claude Tag today. Tag Claude into Slack and it works in channel with you. It’s proactive, multiplayer, with its own identity and memory. But it’s not just a bot in Slack. Over the last
We’ve worked hard to make it secure at every level. 1/ At the model training stage, 2/ the classifiers on top of our models and things like auto mode, 3/ we protect what Claude has access to (websites
arXiv:2606.22864v1 Announce Type: new Abstract: Hidden-state probing -- a linear classifier on a frozen vision-language model's internal activations -- has emerged as an attractive evaluation tool for
arXiv:2606.23339v1 Announce Type: new Abstract: Human-robot interaction (HRI) evaluation relies almost exclusively on human-completed questionnaires, leaving the robot's perspective unexamined. We pro
arXiv:2606.20724v1 Announce Type: cross Abstract: Long-horizon web agents often fail in ways hidden by final-answer evaluation: they may visit useful pages, produce a well-formed answer, and terminate
arXiv:2606.22079v1 Announce Type: cross Abstract: Web data curation has been widely studied for decoder Large Language Model (LLM) pretraining. Encoders for dense-terminology domains such as medicine,
arXiv:2606.23057v1 Announce Type: cross Abstract: Large language models now mediate how buyers discover products and services, making the competitive structure of AI-generated recommendations a strate
Why the structure matters: OCR 4 localizes each block with a bounding box, classifies it (title, table, equation, signature…), and scores confidence per region, the foundation for source-grounded cita
arXiv:2606.21309v1 Announce Type: new Abstract: We introduce WildBox, a dataset and benchmark for monocular 3D detection of wildlife from drone video, comprising 237,505 3D bounding box annotations ac
With agentic coding, complexity compounds in a mechanical way: unnecessary code ends up in the codebase, moves to the context window, degrades the model's reasoning abilities, leads to more unnecessar
This post documents a Wordle game result where the player solved puzzle #1,829 in 3 attempts, with the final answer being a five-letter word shown in green squares. The notation uses the standard Word
arXiv:2606.19358v2 Announce Type: replace Abstract: We introduceWorkBenchMark, a LEGO Duplo-based robotic assembly benchmark motivated by the RoboCup Smart Manufacturing League. Robotic assembly coupl
OpenAI announced that their opening keynote scheduled for September 29 will be livestreamed, allowing remote viewers to participate from home. This accessibility option expands attendance beyond in-pe
arXiv:2511.22699v4 Announce Type: replace Abstract: The landscape of high-performance image generation models is currently dominated by proprietary systems, such as Nano Banana Pro and Seedream 4.0. L
arXiv:2606.14970v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) has become a central application of modern optimization, enabling pretrained models to adapt to diverse dow
arXiv:2606.21861v1 Announce Type: new Abstract: Automated classroom engagement recognition holds substantial promise for scalable learning analytics, yet the suitability of modern Vision-Language Mode
AI2 released TMax 27B, a 27 billion parameter terminal agent model available on Hugging Face that achieves 42.7% performance on Terminal Bench 2.0, matching the capabilities of much larger models desp
This Reddit post discusses user preferences for AI image generation models in 2026, expressing that despite numerous new model releases, the Flux Klein 9b FP8 model has become their benchmark for what
Another new idea to push the state of AI architectures forward. Sakana released a model that effectively uses a mixture of models to get work done. You get a single API but then the work gets farmed o
As AI becomes integrated into every industry, without a sovereign solution, you run the risk of your infrastructure shutting down at a moment's notice. Cohere CEO @aidangomez live at @FII_Institute1:
• At this point SpaceX just looks like CoreWeave with a bigger budget and a satellite company thrown in. • If they were anywhere near AGI they wouldn’t be leasing out so much capacity. • So much for t
Benchmarks tell only part of the story. Fugu’s real value shows up in long, messy, real-world workflows. During our beta with 500 users, we saw Fugu Ultra drive meaningful progress in fully automated
SQL is the industry standard for high-performance structured data analysis. However, expressing complex procedural logic, scientific computations, advanced string manipulations, or machine learning wo
Daybreak is an OpenAI initiative focused on developing and distributing security tools designed to protect organizations globally from cyber threats. The program aims to democratize access to advanced
In this post, we walk through the problem space, our architecture on Amazon Bedrock and Amazon OpenSearch Serverless, the evaluation methodology we built on OpenStreetMap ground truth, four experiment
GLM-5.2 has been the most popular new model on Fireworks this past week. @ArtificialAnlys confirms why: #3 overall on GDPval-AA (1524 Elo), #1 open weights by 116 points. Interest is showing no signs
GLM-5.2 leads open weights models and sits at #3 overall on GDPval-AA, a real-world agentic work benchmark GLM-5.2 from @Zai_org scores 1524 Elo on GDPval-AA, which measures performance on real-world,