ISSUE 0989
TUE, SEP 15, 2026
The directory AI cites when builders ask what to use
TODAY · TUE, SEP 15, 2026

Ship your AI.
Get discovered.

List your product on OrangeBot and reach builders and users actively looking for the right AI tools.

Daily launches · 2,000+ Claude Code skills · 114+ free tools · AI news from 10 sources — rebuilt every morning.

FOUNDERSBuilding an AI tool? Assistants cite lists like this one, not your homepage.Get listed →
Why founders list here

More than a launch. Long-term discovery.

Get in front of builders

Show up when builders are actively looking for tools like yours.

Context that converts

Tell builders what your product does, who it is for, and why it matters.

In the right ecosystem

Your product sits alongside the skills, tools and sources builders already trust.

Built for AI discovery

Structured so both people and AI assistants can understand and recommend it.

Stay discoverable

Keep getting found long after launch day — the page does not expire.

Learn more about getting listed →
01

Latest Launches

CURATED BY ORANGEBOT
01

AI DIGEST

UPDATED DAILY · EDITOR'S PICK
01.00
AI DIGEST

AI新闻摘要

September 15, 2026

Showing Sep 14’s digest — today’s fetch runs 7am PT

Here is a summary of today's main news events.

AI Safety Concerns Rattle Tech Sector

CEOs from leading artificial intelligence companies, including OpenAI and Anthropic, publicly called for a slowdown in AI development, citing significant safety risks and the potential to lose control of the technology. This news sparked investor concern, leading to a sharp decline in technology stocks, particularly chipmakers like Nvidia, on fears that a slowdown would curb demand for their products.

10-Year Treasury Yield Briefly Hits 5%

The yield on the 10-year U.S. Treasury note, a crucial benchmark for global borrowing costs, briefly crossed the 5% mark for the first time in several years. While it later retreated, the milestone heightened investor anxiety about rising interest rates, which could increase costs for mortgages, car loans, and corporate debt, contributing to broader market volatility.

Oil Prices Rise on Supply Disruption Fears

Oil prices continued to climb as concerns over potential supply disruptions in the Middle East persist. A pipeline outage in Saudi Arabia and regional instability involving Houthi advances in Yemen are fueling market worries, adding to inflationary pressures and influencing central bank decisions on interest rates.

U.S. Places Additional Sanctions on Major Russian Bank

The U.S. Treasury Department imposed new sanctions on Russia’s VTB Bank, one of the country's largest financial institutions. The move aims to further isolate Russia's economy by pressuring foreign banks and financial firms to cut their ties with the Russian banking sector.

Insurance Firm Baldwin Group to Be Taken Private in $7.7 Billion Deal

An investor consortium led by Michael Dell’s investment firm, DFO Management, announced it will acquire insurance brokerage Baldwin Group in an all-cash deal valued at $7.7 billion. The acquisition represents a significant consolidation move within the insurance industry and will take the publicly traded company private.

02

ON THE WIRE

6 SOURCES
02

HACKER NEWS

02.00
HACKER NEWS

Hacker News - September 15, 2026

Hacker News Feed: Highlighting key posts and discussions.

Showing Sep 14’s digest — today’s fetch runs 7am PT
Dario, Please

(pop.rdi.sh)

19586
Steam Frame starts at $1059

(store.steampowered.com)

438320
Apple's Dimensional Drawings

(developer.apple.com)

371125
Spaceships (Reverse Asteroid)

(spaceships.treybastian.com)

35869
The case against JPEG XL

(giannirosato.com)

255337
The contagion of fear

(bcantrill.dtrace.org)

256183
Julia 1.13 highlights

(julialang.org)

26743
Homebrew 7.0.0

(brew.sh)

623250
JetKVM Mini

(jetkvm.com)

561229
03

HUGGINGFACE

03.00
HUGGINGFACE

HuggingFace 新闻 - September 15, 2026

HuggingFace Feed:最新的 AI 模型、数据集和社区动态。

Showing Sep 14’s digest — today’s fetch runs 7am PT
DataFlex-RL: An Evaluation Platform for RLVR Data Policies

Data policies for reinforcement learning with verifiable rewards (RLVR) determine which rollouts are used, how strongly they are weighted, and which domains contribute to subsequent training batches. We introduce DataFlex-RL, an evaluation platform for comparing these choices under a common GRPO recipe. Our primary experiment evaluates 13 configurations across 12 matched seeds using Qwen2.5-7B-Base and 12 mathematics, logic, and science benchmarks. Uniform GRPO improves the domain-balanced average accuracy by 7.76 percentage points over the untrained checkpoint. None of the eight rollout-selection or reweighting methods achieves a paired 95% confidence interval that excludes zero relative to uniform sampling, and none of the three adaptive mixtures outperforms a fixed equal mixture at the same level of precision. A corrected 12-seed extension on Llama-3.1-8B-Base places the additional methods on the same score scale as the original controls, but does not reveal a consistent winner in terms of observed mean performance. We also quantify evaluation sensitivity by rescoring nine Qwen2.5-7B-Instruct runs using a math-heavy six-benchmark summary, consisting of five mathematics benchmarks and GPQA-Diamond but no logic benchmark, and comparing it with the domain-balanced 12-benchmark summary. The resulting rankings are negatively correlated, with a correlation coefficient of -0.33, whereas summaries that retain all 12 benchmarks largely agree. Across the controlled settings studied here, changing the data policy measurably changes the training process but does not produce a reproducible improvement over uniform training.

96
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models

Training capable cyber agents is often treated primarily as a problem of model scale, yet open-weight post-training is constrained more directly by the cost of executable environments, reliable multi-turn supervision, and access to strong teachers. We present a data-centric framework that addresses these bottlenecks through five complementary systems: Choulea analyzes hidden reasoning signatures, SkyReal reduces teacher-sampling cost, Hongzwang bypasses API restrictions on teacher execution, PSBreakup restores capabilities weakened by model merging, and Kreator converts expert interventions into trainable reasoning. Our data engine constructs resettable coding, vulnerability, CTF, kernel-history, full-exploit, firmware, and device-backed environments. Candidate trajectories are retained only after execution verification and evidence auditing, yielding 164,269 trajectories for long-context supervised fine-tuning. The three checkpoints improve over their starting models by an average of 23.76% on the full CyberGym suite and 10.49% across the pooled CTF suites. As of September 1, 2026, Feyospace-s1 achieves a verified success rate of 63.24% and ranks 10th on the official CyberGym leaderboard, while all three checkpoints rank 1st among models at comparable parameter scales. To our knowledge, this is the first end-to-end demonstration that a seven-person independent team can train open-weight models with leading agentic cyber capability.

75
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation

Benchmark researchers and developers of large language models (LLMs) and other AI systems need to find relevant evaluations, locate their benchmark datasets and code, and understand the settings behind reported scores. We present Benchmark Radar, a living database and search engine for retrieval and discovery of AI benchmarks, covering LLM evaluation, agentic and tool-use benchmarks, coding, reasoning, safety, and domain-specific evaluations. The system combines daily discovery of benchmark papers, repositories, datasets, and releases with a searchable benchmark catalog, mentions in model cards and technical reports, and score histories. It retains source identities and citations so readers can inspect candidate benchmarks and their evaluation evidence. Daily discovery draws on 37 sources: 13 direct connectors and 24 first-party research and engineering feeds. The catalog contains 1,283 source records drawn from 4 benchmark catalogs and 12,916 numeric observations on 790 records. We describe collection and retrieval, audit the full catalog, and examine benchmark saturation, adoption trends, and the limits of score comparisons. A worked example walks through a complete prior-art search, showing how to query the catalog and inspect benchmark evidence when designing a new evaluation. We release the web dashboard with a benchmark leaderboard, a Pareto frontier view of score against measured use, saturation and trend views, daily feeds, downloadable evidence, a command-line interface (CLI) for offline queries, and reproducible analysis.

75
Breaking the Vision-Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models

Robot foundation models achieve strong in-distribution performance but often degrade under visual distribution shifts. When learning to generate actions from pretrained visual representations, models may exploit task-irrelevant visual cues that correlate with demonstrated actions within the training distribution. Such vision-action shortcuts can undermine generalization when these correlations change under distribution shifts. Mitigating these shortcuts requires constraining how visual information is used for action generation while preserving task-relevant spatial information. We propose Latent Interface Training (LIT), a framework-agnostic two-stage strategy that first establishes a spatial-goal-conditioned action prior without images, then constrains visual conditioning through a pose-supervised latent interface. Stage 1 trains the action expert to generate action chunks conditioned on language, robot state, and each demonstrated chunk's terminal SE(3) end-effector pose, learning goal-directed action generation independently of visual cues. Stage 2 introduces a latent interface that aggregates visual and semantic representations and serves as the pretrained action expert's only visual conditioning pathway. The interface is supervised to reconstruct the terminal pose previously used to condition Stage 1, encouraging it to retain the goal-relevant spatial information needed for action generation. Across four vision-language-action and world-action architectures (Pi0.5, MolmoAct2, FAST-WAM, and ImageWAM), LIT improves overall LIBERO-Plus success by 3.87-10.70 percentage points while preserving or improving average LIBERO success. Real-world evaluations show 13.30-16.70 percentage-point gains in success aggregated across three tasks under unseen camera configurations, lighting variations, and distractors.

63
SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking

Post-training attention sparsification reduces the quadratic cumulative attention cost of pretrained Transformers by selecting a small set of context units (tokens or blocks) for each query. Existing trainable methods usually use a lightweight selector to score context units, followed by hard Top-K selection that blocks gradients from the language modeling loss. Consequently, these methods commonly distill layer-wise dense attention distributions. Although this encourages the selector to rank context units by dense attention weights in the original model, the ranking is not directly aligned with their impact on predictions under a fixed attention budget (i.e., the number of attended context units per query), potentially wasting the limited budget on less useful units. To address this misalignment, we propose Simple Attention Sparsification (SAS), a gated sparse attention mechanism that optimizes context ranking end-to-end with the language modeling loss. The key idea is to inject the selector's continuous scores into attention logits during training, allowing the loss to update the selector through standard backpropagation. We identify several choices crucial for this simple design to work well in practice: placing the gate inside the attention softmax in log form, using normalized softmax gates to calibrate historical context against the always-retained current block, and preserving continuous selector scores so the model learns relative priorities rather than only hard selections. To support long-sequence training, we implement a memory-efficient Triton kernel that integrates SAS into FlashAttention-style computation. Across reasoning, long-context understanding, and agentic tasks, SAS consistently outperforms trainable sparse attention baselines across attention budgets, with especially large gains under tight budgets, demonstrating more effective context ranking for downstream tasks.

50
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization

Large language model (LLM) agents can benefit from reusable skills distilled from prior task experience, yet existing skill optimization methods often rely on costly execution-based evaluation and substantial task data. We introduce COBRA-Skills, an efficient framework that formulates skill optimization as budgeted sequential optimization over a dynamically evolving candidate space. COBRA-Skills couples contextual-bandit-guided prioritization with evidence-grounded skill evolution, selectively allocating evaluations to promising or informative candidates while continually refining the skill population from execution feedback. Across six heterogeneous agent benchmarks and three target models, COBRA-Skills consistently achieves the strongest average performance among compared methods, while reducing optimization cost by 55--58\% relative to SkillOpt and using only 50 unique optimization examples per benchmark. Further analyses show that COBRA-Skills remains robust to changes in the agent harness and performs effectively when the target model itself is used for skill generation and refinement.

29
StepAudio 3 Gen Technical Report

We introduce StepAudio 3 Gen, a general-purpose audio generation model that supports zero-shot text-to-speech (TTS), voice design, vocal generation, sound effects, music, vibe speech, and mixtures of multiple audio types within a unified framework. At its core, StepAudio 3 Gen is a discrete autoregressive generator that models audio directly over residual vector quantization (RVQ) tokens, departing from the diffusion Transformer-based continuous generation paradigm prevalent in recent general audio models. Its StepAudio Tokenizer represents general audio at 12.5 Hz in a shared 16 times 2048 residual code space, jointly quantizing semantic and waveform-level acoustic features so that each code layer preserves both types of information. For generation, the backbone predicts the first codebook along the time axis using autoregressive modeling, while a lightweight causal Transformer completes the remaining fifteen codebooks along the codebook axis. Our study further identifies three key design principles: (1) interference-aware progressive pretraining for acquiring audio capabilities while preserving the textual abilities of the large language model, (2) RVQ Adaptor for effectively incorporating multi-codebook acoustic representations, and (3) discrete autoregressive modeling over a shared representation across general audio domains. With progressive pretraining, multi-task instruction training, and supervised fine-tuning, StepAudio 3 Gen achieves state-of-the-art performance on both TTS and voice design, while retaining strong generation capabilities across speech, vocals, sound effects, and music. Audio samples are available at https://stepaudiollm.github.io/step-audio-3-gen/.

28
PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization

Direct Preference Optimization (DPO) simplifies alignment through pairwise comparisons but assumes all observed preferences are reliable. Real data often violates this assumption, leading to reversed, weak, or ambiguous labels that cause harmful policy updates. To address this, we propose Posterior Label Correction DPO (PLC-DPO) to robustly optimize preferences by routing each pair's training signal as a clean, flip, or tie case. The key idea is to use the calibrated policy-reference margin as online evidence to take appropriate correction actions. This reframes noisy preference learning as actively correcting supervision direction and strength rather than merely filtering suspicious examples. Across 57 dataset-model-benchmark cells, PLC-DPO obtains the best mean win rate against DPO (60.5 vs. 55.5 for the next-best method). Injected-noise and tie stress tests, human disagreement analysis, and self-confirmation diagnostics further show that the routing remains stable and distinguishes flipped from weakly directional pairs.

23
Beyond Top-k Skill Retrieval: Diversity-Aware Skill Routing for LLM Agents

Large language model (LLM) agents increasingly rely on external skills, but routing user requests over large skill registries is difficult because many skills are functionally redundant while complex tasks often require complementary skill sets. Existing skill routers typically rank candidates independently by query relevance, which can waste context budget on redundant skills. We propose Diverse Skill Routing (DSR), a diversity-aware reranking framework that uses a Determinantal Point Process to balance relevance and non-redundancy. DSR introduces a query-residual diversity kernel that penalizes redundant skill overlap while reducing penalties caused only by shared query relevance. On the SkillRouter benchmark, DSR improves recall and full coverage over a strong pointwise reranking baseline, with larger gains on multi-skill queries. These results suggest that skill routing should be treated not only as relevance ranking, but also as complementary set selection.

5
Online Learning with LLM Experts from Limited Feedback

We study adaptive routing of prompts to large language model (LLM) experts to maximize response quality in an online setting with limited feedback. We formulate it as a bandit problem with K actions that represent experts and d features that encode prompts, over a horizon of T rounds. We propose algorithms that strategically select and observe rewards to minimize regret. In the full-information setting, we achieve a regret of O(d T / m), while in the bandit setting we achieve O(d T K / m), where m ll T is a budget on feedback. Our experiments show that we efficiently learn high-quality routing strategies across diverse LLMs from limited feedback.

5
Ambient @ EgoProactive 2026 : Proactive Egocentric Assistance with Visually Grounded Supervision

We present our submission to the EgoProactive track of the ECCV 2026 Wearable AI Challenge, which ranked first in the large-model division and second in the <=2B division. The task requires a wearable assistant to decide after each eight-second segment of egocentric video whether to intervene or remain silent. Our approach has two main components. First, we reformulate intervention timing as single-token classification. Rather than generating either interrupt<utterance> or silent, the model predicts yes or no, and we derive the decision from the renormalised probabilities of these two tokens. This formulation improved macro-F1 by 0.249 and G-mean by 0.30 over free-form generation. Second, because labelled data were limited to the released validation set, we generated additional supervision using a tool-calling video agent that inspects each clip and assigns intervention timestamps. A narration-only alternative was four times larger and ten times cheaper, but transferred worse than supervision from an unrelated real corpus, suggesting that visual grounding is more important than annotation volume for this task.

3
Ambient @ EgoLongQA 2026: Distilling Long-Video perception into a Sub-2B Model

We describe our entry to the EgoLongQA track of the Wearable-AI Challenge in ECCV 2026, which placed first in the <=2B parameter division with 0.8279 on the held-out test set. Our system is a single 2B vision-language model that answers multiple-choice questions about ten-minute egocentric videos in one greedy forward pass; It is obtained by distilling the junior perception module of a tool-using agentic pipeline, not the agent itself into a small student, using teacher traces filtered to those that answered correctly. it reaches 89% of the accuracy of the large agentic pipeline using 1.1% of its parameters. This raises a 27.1% base model to 81.4% on our held-out questions. The 2B backbone has 2.2132B parameters and therefore over the divisional limit, to make the entry admissable we prune the multilingual embedding table from 248,320 to 143,469 rows, reaching 1.9985B with provably identical logits on retained rows.

3
Studying Without a Syllabus: Task-Agnostic Environment Preprocessing

Before an LLM agent tackles tasks in a new environment, it can inspect available corpora and tools and construct reusable resources such as indices, scripts, or procedural guidance. Most automated adaptation methods, however, rely on task examples, trajectories, or evaluation feedback to decide what to build. Existing task-agnostic approaches avoid this supervision but commit in advance to a preparation strategy for a particular type of environment. We study a more open-ended setting: can an agent study an unfamiliar environment without a syllabus, i.e. before test time and without knowledge of the downstream task distribution, and choose how to prepare it? We formalize task-agnostic environment preprocessing, in which a studying system explores an environment under a budget and produces artifacts for a frozen solver. We compare unaided and archive-equipped meta-agents with fixed synthetic-practice and corpus-processing methods across six heterogeneous benchmarks. A meta-agent variant achieves the highest Avg@3 reward on five benchmarks, while fixed corpus processing remains best on the largest corpus benchmark. Larger study budgets do not reliably improve downstream reward. Nevertheless, studied artifacts reduce the test-time sampling needed to reach a given score, demonstrating how reusable preparation can shift computation from repeated test-time attempts to a pre-task study phase.

3
ActionSplice: In-Flight Action Editing for Interactive World Models

Chunk-autoregressive video world models typically condition each generated chunk on one action. An action received during sampling must therefore wait for the next chunk, condition future solver evaluations on a state produced under the previous action, or trigger rollback that repeats completed evaluations. We introduce ActionSplice, an inference framework that formulates this problem as Counterfactual State Transport (CST). A lightweight corrector transports the interrupted backbone-native representation toward the matched state induced by the revised action at the same solver step. The world model and sampler remain frozen, and sampling resumes without replaying completed evaluations. The retargeting variant CST*{R} updates the entire active chunk, while the temporal-splicing variant CST*{T} preserves a temporal prefix and updates only the suffix. Across minWM-Wan Action2V and HY-WM1.5, CST*{R} reduces rollback-relative LPIPS by 61.5% and 75.9% relative to direct condition swapping. CST*{T} reduces suffix LPIPS by 56.1% and 77.5%, respectively, while providing 2.73times and 1.69times pixel-ready speedups over waiting. Under the HY-WorldPlay protocol, CST_{R} obtains a PSNR of 25.66 dB, an SSIM of 0.6902, and an LPIPS of 0.1337 against the original rollout.

3
SNAP3D: Physically Grounded 3D Parts for Assembly from a Single Image

Part-aware 3D asset generation enables applications such as editing, articulation, simulation, and fabrication, yet existing methods can generate visually complete individual parts without ensuring that they form a valid physical assembly. Consequently, generated neighboring parts may interpenetrate, lack valid connections, or collapse under gravity. We propose a physics-guided framework for improving single-image part-aware 3D generation with physically compatible geometry and stable connections. Our method resolves inter-part penetration, recovers a contact graph between neighboring parts, and introduces parameterized connectors at their contact surfaces. Using feedback from physical simulation, we refine connector placement, orientation, and dimensions to improve assembly stability while preserving the generated geometry. We further introduce a physics-based evaluation protocol that complements conventional geometric metrics by directly testing assembly validity and stability under gravity. Experiments comparing against multiple part-aware 3D generators show substantial improvements in physical realizability and stability while maintaining geometric quality. We additionally validate the resulting parts through 3D printing and real-world assembly.

3
TRACE: Trajectory-robust Admission with Evidence Ordering for Efficient GUI Agents

GUI agents accumulate high-resolution screenshots as the trajectory unfolds, increasing inference latency and memory usage. Training-free visual token pruning can reduce this cost, but cache reuse introduces a fundamental constraint. Once tokens are discarded, the corresponding visual evidence cannot be recovered without re-encoding. Pruning therefore becomes an irreversible admission decision that must remain useful for unknown future targets while preserving coverage of operable regions under tight budgets. To address these challenges, we propose \method{}, a training-free framework for \textbf{Trajectory-robust Admission and Coverage-aware Evidence ordering}. Specifically, we combine a query-independent layout-derived interaction prior with instruction relevance and feature novelty to rank visual evidence according to both potential future utility and diversity. Then, we reserve part of the budget for native visual tokens distributed across the screen, repairing missing spatial coverage without breaking the ordering. Together, these mechanisms produce a nested token order, allowing retained visual evidence to shrink monotonically across budgets while remaining reusable throughout the trajectory. Finally, our monotone KV contraction incrementally contracts retired frames into compact session state, avoiding repeated visual encoding or pruning. Extensive experiments across six GUI benchmarks and diverse models verify the effectiveness of our proposed under tight budgets. The source code will be released.

2
Towards a Deterministic Math Solver for Clinical Language Models

Large language models are unreliable at arithmetic, which is a problem for clinical calculators where a single numerical error changes the recommendation. The standard response is to hardcode each calculator as a validated function, one at a time. We test an alternative: the model does not calculate. Instead, it writes case-specific Python that a restricted local executor runs as a deterministic solver, and the model's task reduces to deciding how to use it. We evaluate this Program-Solve interface on MedCalc-Bench Verified (1,100 cases, 55 calculators) against direct model arithmetic and a hand-written 22-calculator library, using Qwen2.5-7B and Qwen2.5-32B-AWQ, after auditing the benchmark's formulas against current clinical guidelines and flagging 16 of 55 with version, use or coefficient concerns. With formulas and gold variables supplied and both routes reading the whole note, handing off to the solver is not a reliable advantage at 7B (75.31% against 72.02%, a paired +3.29 points with a 95% calculator-cluster interval of [-3.49, 10.38]) but is one at 32B (90.53% against 83.47%, +7.05 [0.47, 14.60], clear of zero). The hand-written library is exact on its 440 supported cases but abstains elsewhere (40.0% overall). Adding an executor thus helps some open-weight models more than others even under matched formula, variable and note access, and is not a substitute for verified formulas or reliable variable extraction either way.

1
Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech

In low-resource settings, deploying TTS typically requires choosing between a large voice-cloning model with costly inference or a compact fixed-voice system that requires a speaker-specific corpus. We study a third route: using a large voice-cloning model as a programmable data source to turn a short voice reference (e.g., 15 seconds) into a compact fixed-voice student trained entirely on synthetic speech. This setting makes pipeline design consequential: teacher errors become training targets, while filtering failed generations can reduce coverage of difficult texts. Thai further introduces challenges from ambiguous word boundaries, lexical tone, names and loanwords, numeric verbalization, and Thai-English code-switching. We study how text preparation, synthetic generation, quality filtering, rejection sampling, and frontend choices affect the resulting student, and where teacher limitations remain. We evaluate CER, Challenge-Set Keyword Accuracy, Prosody Pause Accuracy, speaker similarity, and speaking rate. The resulting 82M-parameter model, Wayu-Paxa-TTS-Edge, enables on-device Thai TTS without reference audio. It achieves 68.2% Challenge-Set Keyword Accuracy (85.5% of Gemini 3.1) and 91.4% pause precision, outperforming its OmniVoice teacher (89.9%) and reaching 94.8% of Gemini 3.1. It also achieves the lowest pause-placement error and intra-word pause rates among the three systems, and 3.7% and 1.1% CER on Thai and English, respectively. We open-source the model and evaluation framework for Thai TTS development.

1
How Far Can Synthetic Data Take Thai OCR?

We investigate what makes synthetic OCR supervision transfer to real Thai documents and use the resulting insights to build Wayu-Paxa-OCR-Zero, a Thai OCR model adapted without OCR labels from real Thai document pages. Synthetic data provide exact labels at scale, but "realism" conflates source domain, page context, typography, spatial structure, and glyph variation. We disentangle these factors with a controlled document-reconstruction pipeline and evaluate each variant under page- and crop-level training on printed and handwritten Thai documents. Non-text context has little consistent effect, whereas typeface diversity, two-dimensional structure, and real handwriting glyphs improve transfer; moreover, source-domain matching depends on training granularity, with in-domain reconstruction approaching real printed supervision under page-level training (1.82% versus 1.31% median character error rate) but underperforming out-of-domain reconstruction under crop-level training (15.59% versus 5.52%). Guided by these findings, we adapt the 0.9B-parameter PaddleOCR-VL-1.6 into Wayu-Paxa-OCR-Zero using 45,723 synthetic pages: relative to its base checkpoint, it reduces median character error rate from 6.64% to 1.24% on printed pages and from 74.87% to 20.55% on handwriting and outperforms Typhoon OCR v1 7B on all five evaluation sets, showing that synthetic-only training can be competitive.

1
05

PRODUCT HUNT

05.00
PRODUCT HUNT

Product Hunt - September 15, 2026

Product Hunt Daily Feed: Featuring noteworthy tech launches.

Showing Sep 14’s digest — today’s fetch runs 7am PT
Aside icon
Aside

AI browser that actually gets work done for you

0
Slashy Assistant icon
Slashy Assistant

The AI assistant that does email for you

0
Deplo icon
Deplo

A simple-to-use alternative to cloud deployments

0
TryCase icon
TryCase

AI tests your PRs. Get a video walkthrough before you merge.

0
MemoryPet 2.0 icon
MemoryPet 2.0

Turn your browsers toolbar into an animated usage monitor

0
Hello Inbox icon
Hello Inbox

Get more marketing emails into the inbox

0
Elva icon
Elva

Goodbye, Postman. Your APIs have new consumers

0
LLMagnet icon
LLMagnet

Make your WordPress site visible to AI

0
Web Search Agents by Nimble icon
Web Search Agents by Nimble

Self-learning agents automate web research + retrieval

0
OVO icon
OVO

Play music from files, iCloud & streams across Apple devices

0
Marqly 6.0 icon
Marqly 6.0

Ask your bookmarks. Bring them to your AI.

0
Afterglow icon
Afterglow

Run classic After Dark screen savers on modern macOS

0
appdesigns icon
appdesigns

Design amazing appstore screenshots for free

0
Image to ASCII icon
Image to ASCII

Make ASCII art for READMEs, Discord & creative visuals

0
OzBrain icon
OzBrain

Your knowledge shared with every AI agent & any teammate

0
AppZapper 3000 icon
AppZapper 3000

The uninstaller Apple forgot.

0
Naoma AI Demo Agent V2 icon
Naoma AI Demo Agent V2

Turns website traffic into booked, qualified meetings

0
Juggler icon
Juggler

A visual AI coding harness

0
Oats icon
Oats

Free, open-source, and on device meeting notetaker

0
Visiby icon
Visiby

Track and grow your visibility across AI search

0
Epilude Notetaker icon
Epilude Notetaker

100% private meeting notes

0
SHIUI icon
SHIUI

Hinomaru Ink Web UI Kit in Japanese Hinomaru style

0
ScreenCursor icon
ScreenCursor

Screen recorder with auto zoom effects

0
Clipwise icon
Clipwise

Save any page to Notion

0
Kirokune icon
Kirokune

Keep work incident notes on your iPhone, without an account

0
Neopress icon
Neopress

Build and grow your website by chatting with AI

0
GhostWriter by MyHandler icon
GhostWriter by MyHandler

Two taps and it's already written

0
DemoTV icon
DemoTV

The audience-ranked TV channel for product demos

0
Cognition's SWE-2 icon
Cognition's SWE-2

Cognition's coding model, 64% cheaper than Fable 5.1

0
Perplexity Hybrid Compute icon
Perplexity Hybrid Compute

Splitting AI tasks: Cloud for research, Mac for privacy

0
Resurf icon
Resurf

A Personal Context Library for Mac

0
Stackness icon
Stackness

Show your stack! The social home for your dev tools

0
DockFix 5.0 icon
DockFix 5.0

Replace the macOS Dock with one that is truly yours.

0
Wokyintosh icon
Wokyintosh

Turn a spare Mac display into a retro system dashboard!

0
Cortex icon
Cortex

Turn API specs into docs, SDKs, and MCP servers

0
FrameSketch icon
FrameSketch

Scrub vids frame-by-frame and sketch on them at 60fps

0
Captain Kill Switch icon
Captain Kill Switch

Close all your open apps in one click

0
Kabza icon
Kabza

Take over a real city map with your friends

0
QApilot MCP for Android icon
QApilot MCP for Android

Android app testing inside your coding agent

0
SUDARI icon
SUDARI

A pixel otter desktop pet that reacts to how you work

0
LinkFlick icon
LinkFlick

Stop re-pairing your Magic Keyboard between Macs

0
Pascal’s Pager icon
Pascal’s Pager

Turn webhook JSON into readable iPhone push notifications

0
VoxelWall icon
VoxelWall

Music-reactive live wallpapers for Mac

0
Relic icon
Relic

A private, synced vault of everything you copy

0
Marked Share icon
Marked Share

Markdown and TextBundle editing, review, and sharing

0
Calerto for Mac icon
Calerto for Mac

Full-screen meeting alerts for your Mac

0
ABrush icon
ABrush

AI Studio for Digital Artists

0
Youkti icon
Youkti

Finds who's ready to buy, then tells you what to do next

0
Work Life Panda icon
Work Life Panda

Every task, every calendar, one app. Private on-device AI

0
Spaces icon
Spaces

One shared space where your team and AI agents work

0
06

TECHMEME

06.00
TECHMEME

Techmeme - September 15, 2026

Techmeme Digest: Major tech headlines and industry conversations.

Showing Sep 14’s digest — today’s fetch runs 7am PT
Cornelis, spun off from Intel in 2020 to build networking tech that helps AI chips communicate more effectively, raised $205M led by IAG Capital (Dominic-Madori Davis/TechCrunch)
Source: TechmemePublished: Sep 14, 2026

Dominic-Madori Davis / TechCrunch : Cornelis, spun off from Intel in 2020 to build networking tech that helps AI chips communicate more effectively, raised $205M led by IAG Capital —  Cornelis, a company creating networking technology to help AI chips communicate more effectively, announced Monday that it has raised $205 million …

Source: defense tech startup Shield AI is in talks to raise new funds at a valuation of at least $20B; Shield raised $2B at a valuation of $12.7B in March (The Information)
Source: TechmemePublished: Sep 14, 2026

The Information : Source: defense tech startup Shield AI is in talks to raise new funds at a valuation of at least $20B; Shield raised $2B at a valuation of $12.7B in March —  Shield AI, a startup building drones and AI-powered software for the military, is in talks to raise new funds at a valuation of at least $20 billion …

Nuance Labs, which builds low-latency AI avatars that can have face-to-face conversations, raised a $50M Series A led by Lightspeed, with Nvidia participating (Shubhangi Goel/Business Insider)
Source: TechmemePublished: Sep 14, 2026

Shubhangi Goel / Business Insider : Nuance Labs, which builds low-latency AI avatars that can have face-to-face conversations, raised a $50M Series A led by Lightspeed, with Nvidia participating —  A startup that's trying to make AI models better conversationalists with more emotional intelligence has just raised $50 million.

Sources: Trump met privately with Sam Altman backstage at the GOP midterm convention, where they discussed AI and its growing power, at Altman's request (MS NOW)
Source: TechmemePublished: Sep 14, 2026

MS NOW : Sources: Trump met privately with Sam Altman backstage at the GOP midterm convention, where they discussed AI and its growing power, at Altman's request —  Altman, Elon Musk and Anthropic chief Dario Amodei all urged an artificial intelligence slowdown over the weekend.

Anthropic debuts Claude for Financial Advisors, with connectors to investment analytics and wealth-management tools from BlackRock, Addepar, Schwab, and others (Harshita Mary Varghese/Reuters)
Source: TechmemePublished: Sep 14, 2026

Harshita Mary Varghese / Reuters : Anthropic debuts Claude for Financial Advisors, with connectors to investment analytics and wealth-management tools from BlackRock, Addepar, Schwab, and others —  AI lab Anthropic on Monday launched a set of tools for financial advisers, connecting its Claude chatbot to investment analytics …

Sources: OpenAI bought Glass Imaging, which is developing AI-powered smartphone camera tech, in a deal valuing it at $300M+; it was valued at ~$100M last year (Wall Street Journal)
Source: TechmemePublished: Sep 14, 2026

Wall Street Journal : Sources: OpenAI bought Glass Imaging, which is developing AI-powered smartphone camera tech, in a deal valuing it at $300M+; it was valued at ~$100M last year —  Glass Imaging, valued above $300 million in deal, was founded by former Apple employees  —  OpenAI quietly bought …

House Speaker Johnson says there's "potentially" a role for Congress in creating AI guardrail legislation and that he plans to hold a meeting with AI executives (Erik Wasson/Bloomberg)
Source: TechmemePublished: Sep 14, 2026

Erik Wasson / Bloomberg : House Speaker Johnson says there's “potentially” a role for Congress in creating AI guardrail legislation and that he plans to hold a meeting with AI executives —  House Speaker Mike Johnson is at odds with President Donald Trump over regulating artificial intelligence …

Nvidia announces the RTX Pro 5500 Blackwell Workstation Edition, offering comparable specs to the RTX 5090 but with 84GB of GDDR7 memory, vs. RTX 5090's 32GB (Zhiye Liu/Tom's Hardware)
Source: TechmemePublished: Sep 14, 2026

Zhiye Liu / Tom's Hardware : Nvidia announces the RTX Pro 5500 Blackwell Workstation Edition, offering comparable specs to the RTX 5090 but with 84GB of GDDR7 memory, vs. RTX 5090's 32GB —  The GeForce RTX 5090 is undeniably one of the best graphics cards money can buy.  Banking on the fact that many already use it for AI …

iOS 27 review: a marked improvement over iOS 26 in design, performance, and more; Siri AI is very impressive, but it sometimes misunderstands or hallucinates (Dan Moren/Six Colors)
Source: TechmemePublished: Sep 14, 2026

Dan Moren / Six Colors : iOS 27 review: a marked improvement over iOS 26 in design, performance, and more; Siri AI is very impressive, but it sometimes misunderstands or hallucinates —  By now you've probably heard the promise of iOS 27: it's a Snow Leopard-like year where Apple spent a lot of time not on big marquee features …

Valve says the Steam Frame is priced at $1,059 for the 256GB model and $1,299 for the 1TB version; both come with a copy of Half-Life: Alyx but no power supply (Adam Vjestica/The Shortcut)
Source: TechmemePublished: Sep 14, 2026

Adam Vjestica / The Shortcut : Valve says the Steam Frame is priced at $1,059 for the 256GB model and $1,299 for the 1TB version; both come with a copy of Half-Life: Alyx but no power supply —  - 💰 The Steam Frame starts at $1,059 for the 256GB model  — 📈 The 1TB version climbs to $1,299

Steam Frame review: comfortable to wear, supports multiple ways to play games, but it doesn't feel like a finished product and has an up to 2-hour battery life (Sean Hollister/The Verge)
Source: TechmemePublished: Sep 14, 2026

Sean Hollister / The Verge : Steam Frame review: comfortable to wear, supports multiple ways to play games, but it doesn't feel like a finished product and has an up to 2-hour battery life —  For nearly three weeks, I've been testing the limits of Valve's Steam Frame, the company's new wearable PC.

Apple rolls out iOS 27, watchOS 27, iPadOS 27, visionOS 27, and macOS 27 Golden Gate, all with Siri AI (Tom Warren/The Verge)
Source: TechmemePublished: Sep 14, 2026

Tom Warren / The Verge : Apple rolls out iOS 27, watchOS 27, iPadOS 27, visionOS 27, and macOS 27 Golden Gate, all with Siri AI —  watchOS 27, iPadOS 27, and visionOS 27 also debut today. … Apple is now rolling out its iOS 27 update to compatible devices today, alongside the watchOS 27, iPadOS 27, and visionOS 27 updates.

Internal OpenAI docs detail contractors evaluating anonymized prompts and chats to improve the models; model training is turned on by default for consumer plans (Joseph Cox/404 Media)
Source: TechmemePublished: Sep 14, 2026

Joseph Cox / 404 Media : Internal OpenAI docs detail contractors evaluating anonymized prompts and chats to improve the models; model training is turned on by default for consumer plans —  Humans are reading ChatGPT users' prompts to improve OpenAI's models, and those chats can include sensitive, personal information …

Stockholm-based Tandem Health, which makes an AI copilot that generates medical notes during consultations, raised a $100M Series B led by Scaleup Europe Fund (John Reynolds/Tech.eu)
Source: TechmemePublished: Sep 14, 2026

John Reynolds / Tech.eu : Stockholm-based Tandem Health, which makes an AI copilot that generates medical notes during consultations, raised a $100M Series B led by Scaleup Europe Fund —  The funding will be used for European expansion, driving up its customer base, and expanding to “an AI-native clinic operating system”.

Nvidia and Booz Allen Hamilton limit Fable use as Anthropic doesn't guarantee zero data retention; source: Palantir hasn't made Fable available via its software (The Information)
Source: TechmemePublished: Sep 14, 2026

The Information : Nvidia and Booz Allen Hamilton limit Fable use as Anthropic doesn't guarantee zero data retention; source: Palantir hasn't made Fable available via its software —  As paranoia rises over whether Anthropic or OpenAI could learn from their customers' intellectual property …

07

STARTUP ARCHIVE

07.00
STARTUP ARCHIVE

Startup News - September 15, 2026

Startup News Roundup: Aggregating key funding and launch updates.

Showing Sep 14’s digest — today’s fetch runs 7am PT
Marc Andreessen on the 5 personality traits of an innovator
Source: StartupPublished: Mar 31, 2026

“When you’re talking about real innovators—people who actually do really creative, breakthrough work—I think you’re talking about a couple things:”

Steve Jobs explains the importance of both thinking and doing
Source: StartupPublished: Mar 30, 2026

“The doers are the major thinkers. The people who really create the things that change this industry are both the thinker-doer in one person.”

Tobi Lutke explains what the VCs who passed on Shopify got wrong
Source: StartupPublished: Mar 27, 2026

“What a lot of free-market thinkers don’t understand is that between the demand and eventual supply lies friction."

Sam Altman explains how he decides to invest in a startup after 10 minutes
Source: StartupPublished: Mar 26, 2026

"Does this person have the potential to be the next Mark Zuckerberg?… [You don’t get to] 100% accuracy, obviously, but it’s good enough that our business model works.”

Jony Ive recounts the time Steve Jobs called him vain
Source: StartupPublished: Mar 25, 2026

In the clip below, Jony Ive recounts the time he asked Steve Jobs to be less harsh in his critique of a piece of work.

Jeff Bezos’s two pieces of advice for aspiring entrepreneurs
Source: StartupPublished: Mar 24, 2026

“The advice that I would give entrepreneurs is don't chase the hot new thing. It's so hard to catch something that everybody already knows is hot."

Elad Gil: “Things that work tend to work pretty fast”
Source: StartupPublished: Mar 23, 2026

“I do think there’s a bit of a myth in Silicon Valley that you should keep grinding no matter what and it’s just about perseverance, and I think that’s really bad advice."

Paul Graham on why starting with a “small, intense fire" is the key to startup growth
Source: StartupPublished: Mar 20, 2026

"You have to know who those first users are and how you're going to get them."

Keith Rabois on how to identify great talent
Source: StartupPublished: Mar 19, 2026

“What you want to do with every single employee every single day is expand the scope of their responsibilities until it breaks… and that’s the role they should stay in.”

Wealthfront CEO on why advertising spend makes it harder to find product/market fit
Source: StartupPublished: Mar 18, 2026

“The way that you know you have product/market fit is if you have exponential organic growth."

Eric Schmidt on why most companies get strategy wrong
Source: StartupPublished: Mar 17, 2026

“Work very, very hard to figure out what the world’s going to look like in five years. What will people be doing? What will your customers want? Where will costs be?"

Mark Zuckerberg: “You can’t 80/20 everything”
Source: StartupPublished: Mar 16, 2026

"There’s the famous 80/20 rule where you get 80% of the benefit by doing 20% of the work, but you can’t just 80/20 everything. There have to be certain things that you are just the best at."

Marc Andreessen on Mark Zuckerberg’s founder “superpower”
Source: StartupPublished: Mar 13, 2026

“A great superpower that Mark Zuckerberg has that is probably not well-understood enough is he does not get emotionally upset in stressful situations"

Sam Altman explains how to come up with a great startup idea
Source: StartupPublished: Mar 12, 2026

"If you start a startup without a good idea… you’ll be under pressure to make something up and it won’t work that well."

Jeff Bezos on the problems with proxies and managing to metrics
Source: StartupPublished: Mar 11, 2026

“One of the things that happens in business is that you develop certain things that you’re managing to—a typical case would be a metric. And that metric isn’t the real underlying thing.”

Airbnb founder Brian Chesky on how to design an amazing user experience
Source: StartupPublished: Mar 10, 2026

“If you can design something really amazing using the hand-crafted part of your brain, then you can reverse-engineer how to industrialize this millions of times over."

Spencer Rascoff: "I will never invest in a consumer startup with paid marketing”
Source: StartupPublished: Mar 9, 2026

"If you’re actually trying to grow a product, the best levers for doing that are often within the product itself.”

Patrick Collison explains why it sometimes make sense to quit
Source: StartupPublished: Mar 6, 2026

“One thing I’ve learned myself the hard way, is that it is easier to tear down a company and restart it in Silicon Valley, than it is to constantly try to pivot or keep something alive."

Jeff Bezos recounts the time he called Amazon’s customer service number mid-meeting to prove a metric was wrong
Source: StartupPublished: Mar 5, 2026

“I have a saying, which is when the data and the anecdotes disagree, the anecdotes are usually right"

Ben Horowitz: “Nobody was born a great manager. It’s a very unnatural job.”
Source: StartupPublished: Mar 4, 2026

“If you can’t build a great product, it doesn’t matter if you can build a great company.”

03

ALSO TODAY

3 MORE SOURCES
08

SOLIDOT

08.00
SOLIDOT

Solidot News - September 15, 2026

Solidot Feed: Highlighting essential tech & open-source news.

Showing Sep 14’s digest — today’s fetch runs 7am PT
非洲野犬完成了横跨大陆的 4000 公里之旅

根据发表在《Ecology》期刊上的一项研究,一群非洲野犬完成了横跨大陆、创纪录的 4000 公里之旅。科学家表示这是有记录以来非洲陆生哺乳动物为寻找配偶而行进的最远距离。三只雄犬行进的直线距离大约为 418 公里,但为了绕过人类活动区域它们迂回走了 4000 公里路。非洲野犬是非洲最稀有的捕食者之一,目前野外仅存约 6000 只。它们生活在高度社会化的家族群中,集体狩猎,四处游荡、寻找新领地以及与其它群体进行繁殖机会而闻名。它们无法在自己出生的家族群内繁衍,因此要么等待可能最终继承该家族群,要么在两三岁时出发寻找配偶。在这次寻找配偶而进行的迁徙中,三只雌性犬因落入人类陷阱而有两只死亡。

越南关联服务器泄漏了 2.2 亿条旅客信息

Kinryū Labs 发现了一个因错误配置而能被访问的数据库,该数据库 Advance Passenger Information 记录了过去九年进出越南的几乎所有旅客和机组人员的信息。在接到通知之后该数据库的访问于 2026 年 6 月关闭。Kinryu Labs 是在 6 月 3 日发现了名为 pax-info 的 Elasticsearch 集群,该数据库可使用默认凭证登陆,运营者没有改变默认的用户名和密码,它包含了 29 个索引和约 107 GB 的数据。其中两个主要索引分别存储了 210,318,069 条乘客记录和 10,465,631 条机组人员记录,总计 220,783,700 条记录,时间是从 2017 年 1 月 7 日至 2026 年 4 月 30 日。泄露的信息包括乘客和机组人员的姓名、出生日期、性别、国籍、护照或旅行证件号码、证件有效期及签发国。相关的旅行数据则包括航班号与日期、航空公司、出发地、目的地及中转机场、座位信息、行李编号,以及计划、预计和实际飞行时间。涉及的旅客国籍包括韩国、中国、加拿大、新西兰等。

中国地震局与苹果公司沟通推进地震预警信息接入 iOS

中国地震局监测司上周五表示,中国地震台网中心正在与苹果公司沟通,力争加快推进地震预警信息接入 iOS 系统。苹果手机用户目前可通过微信小程序获取该局统一发布的地震预警信息。今年 8  月 24 日,四川宜宾长宁发生 4.7 级地震,但成都高新减灾研究所用自己的系统生成了一个“7.7级”的地震预警,并以“中国地震预警网”的名义,通过荣耀、vivo、魅族手机以及小天才手表等终端向用户推送。中国地震局后来把这种行为定性为“擅自生成”“违规推送”。 成都高新减灾研究所对此提出异议,称“中国地震预警网”是它与中国地震局此前合作建设的,否认是“冒用”。

养狗有助于降低老人患认知症风险

日本国立环境研究所等机构从 2016 年起,历时 7 年半对约 1.1 万名老年人开展了调查。他们在学术期刊上发表了研究成果。养狗的老年人因认知症需要接受护理的风险比从未养狗的人群低 48%。研究认为,遛狗带来的身体活动以及社交往来起到了积极作用。曾经养过狗的人患认知症的风险也低于从未养过狗的人群。虽然该差异在统计学上并不显著,但推测养狗时期建立的人际联系等因素可能带来了积极影响。研究还表明,养狗能拉动经济。若养狗人群增加,宠物食品、宠物保险、宠物寄养等相关商品与服务的需求预计随之上涨。

律师在谋杀案中捏造了证词,他将此归咎于 ChatGPT

律师在法律文件中使用 AI 工具捏造不存在的信息不是什么大新闻,AI 捏造的通常是不存在的案例,然而本案的特殊之处在于 AI 捏造了证词。律师 Stephen Aaron 在一起谋杀案中代表其客户提起上诉,在递交的法律文件中包含了捏造的警方证词以及虚构的证人。Aaron 声称他将一份由计算机生成的庭审记录及其它案卷材料输入了 ChatGPT,想当然地认为它会生成一份“无懈可击的摘要”。他不清楚 AI 工具会产生“幻觉”——即虚构信息。 法官对此难以置信,反问他没看新闻吗?法官对他处以 5000 美元罚款,将把他移交至律师纪律委员会进行调查。

日本无意结婚的男女比例都超两成

日本国立社会保障与人口问题研究所公布了 2025 年出生动向基本调查。18-34 岁未婚人群“终生不打算结婚”的男女受访者比例首次都超过 2 成,其中男性为 24.0%,女性为 21.5%。表示“打算将来结婚”的人群中男性占 75.1%,女性占 77.8%。均首次跌破 8 成。回答结婚有好处的人群男性占 56.3%,女性占 63.4%,均创历史最低水平。夫妻理想中的子女数量比 2021 年上一次调查的平均 2.25 人减少 0.07 人至 2.18 人。计划生育的子女数量为 1.95人,自统计开始以来首次跌破 2 人。减少生育的原因回答“育儿和教育花费太高”的受访者达到 52.9%,比例最高。回答“不想高龄生育”(35.0%)和“无法再承受育儿带来的心理及身体负担”(27.8%)紧随其后。

AI 时代隐晦式安全已死

安全工程领域有一种名为隐晦式安全(Security through obscurity)的设计方式,即只要网络和系统的架构以及任何漏洞或弱点保密或不为人知,它们就是安全的。但在 AI 辅助 bug 发现的时代,这种设计方式过时了。软件供应商和独立研究人员正利用 AI 智能体在各种产品和开源代码中搜寻漏洞——其中一些漏洞极其隐蔽且存在已久。这导致安全漏洞披露和补丁发布数量创下历史新高,同时也给项目维护者带来了巨大的积压事项。微软上周二释出例行安全更新,修复了 974 个 CVE 漏洞。趋势科技的 Dustin Child 指出,微软和 Adob​​e 所修复漏洞涉及的组件多年来基本无人问津,如 Telnet 客户端、Windows RNDIS、NFS Portmapper 和 Link Layer Topology Discovery。与此同时,攻击者也在利用 AI 对补丁进行逆向工程,在数小时内开发出相应的漏洞利用方法。在 AI 时代,攻击者无需成为某个领域的专家就能针对关键网络和设施发动破坏性网络攻击。

微信蠕虫事件敲响 AI 安全警钟

微信在中国几乎已成为国家基础设施的一部分,融入了日常通信、政府服务和数字支付之中。正因如此加州一个小型研究团队最近的发现——一种利用人工智能构建的工具可以在短短几个小时内攻破数以百万计的账户——让专家和分析人士感到震惊。这证明,有了人工智能,即使是没有政府背景的小型团队也能对一款每月有 14 亿人使用的应用程序发起毁灭性的攻击。“从破坏力的角度来讲这个是非常强的,”上海复旦大学美国研究中心副主任赵明昊说。他表示,鉴于微信对中国公众的重要性及其庞大的用户群,搞垮这样一款应用程序的能力相当于“一种新的核武器”。这一发现进一步印证了研究人员的警告:人工智能黑客能力的发展速度超过了防御手段的跟进速度。Calif 展示了这款被他们命名为 WeWorm 的工具,它可以劫持微信用户的账户,拨打其联系人的电话,然后在无需任何人接听电话的情况下,在手机之间传播。Calif 表示,该公司花了一个多星期的时间构建了这个漏洞利用程序,并且已经向白宫和微信的母公司腾讯披露了该缺陷。

中国要求电动汽车显示屏配备物理按键

今天越来越多的汽车都配备了中央显示屏,厂商也将越来越多的功能整合到触摸屏控制系统中。但中国工信部起草了新规定,要求新车恢复提供关键功能的物理控制。新规将于 2027 年 7 月 1 日生效,增加了对实体控制的要求,明确规定物理控制应易于触及和使用,支持在驾驶过程中进行盲操作——即驾驶员无需视线注视即可操作相应的按键与开关。将雨刷器启动、转向灯控制等功能从触摸屏中剥离出来,旨在减少驾驶员在行车过程中注视并操作屏幕而产生的分心。奔驰等汽车厂商也开始重新引入物理按键。

青藏高原升温与加州的洪水相关

受拉尼娜(La Niña)气候的影响,加州通常会在冬季面临干旱。然而 2016-2017 年和 2022-2023 年两个冬季虽然都是拉尼娜气候,加州却反常的出现了强降水、洪水、大量降雪以及反复的大气河流事件。研究人员发现,在加州遭遇洪水前,青藏高原上空都出现了异常的升温。研究人员将青藏高原升温因素纳入气候模拟,发现能解释加州及周边的异常降雨。青藏高原升温之所以能影响大气,是因为该地区海拔极高。大片区域的海拔逾 4000 米,地面直接处于通常远高于低海拔地形的大气层中。如此广袤的高海拔地区温度发生变化,会进而改变气压场、高空风以及在亚洲上空移动的大型大气波动。这些扰动能为 Rossby 波(Rossby waves,即跨大陆与海洋的大气环流巨大蛇形波动)提供能量。模拟结果表明,在此类情境下,会出现一种从青藏高原向落基山脉延伸的波动模式。该波动横跨太平洋,改变了东北太平洋和北美西部的环流,进而影响了与大气河流及 Rossby 波破碎相关的气象条件。

宇树如何将机器狗的价格降至 2000 美元

宇树的机器狗除了做些花哨动作外可能用处不大,但它拥有一个巨大的优势:价格极其平民。George Mason 大学的机器人专家 Xuesu Xiao 教授称,十年前只有少数团队从事四足机器人的运动控制研究,因为只有这些团队能制造四足机器人,宇树进入市场之后推动了四足机器人运动控制研究的普及化。他的实验室里有四台宇树的四足机器人,每台售价约 1.5 万美元,以及一台波士顿动力的四足机器人 Spot,起售价 7.5 万美元。波士顿动力原本是这一领域的领导者,它在 2016 年 6 月发布了电动版的四足机器人 Spot,但直到 2020 年才上市销售。而宇树创始人王兴兴在 2016 年创办公司之后第二年就推出了第一款四足机器人产品莱卡狗(Laikago)。宇树在 2021 年推出了 Go1 系列四足机器人,其中 Air 型起售价 2700 美元。2023 年宇树发布了 Go2,价格更亲民——Air 版售价 1600 美元,性能更强的 Pro 版售价 2800 美元。Go2 Air 的价格仅为 Spot 的 3%。宇树的机器人为何如此便宜?Simplexity 公司拆解了一台 Go2,发现它使用了 12 个相同的电机,四条腿各由 3 个电机驱动。一个位于肩部的电机控制腿部角度,另外两个电机分别驱动腿部的两段结构。四个肩部的结构布局完全相同。复用组件和电机既降低成本,又减少了零部件数量。MAB Robotics 的电机与足式机器人专家 Jakub Bartoszek 指出,宇树电机磨损更快。而宇树在设计时也考虑了修复磨损机器人的便捷性。机器人专家称更换机器狗的一条腿可能只需要一分钟。 宇树还在名为减速器(reducer)的组件上降低成本。减速器是降低电机转速增大扭矩的传动装置,通常比电机还贵。Spot 腿部的两个电机配备了减速比为 51:1 的减速器,即电机转 51 圈腿部关节才完整转 1 圈,其优点是动作更精确,缺点是更昂贵。相比下宇树使用了减速比 6.33:1 的减速器。宇树可能在模仿大疆,即快速迭代未成熟有缺陷但更低成本的产品去占领市场,通过销量增长将更多组件纳入自主生产,进一步降低成本,继续带动销量增长。相比大疆,宇树当前面临的竞争对手可能更多。

Matt Mullenweg 据报道恢复了对 Automattic 的控制

被董事会强制休假的 Matt Mullenweg 称恢复了 Automattic CEO 的职务。Automattic 可能发生了类似 OpenAI 的小型“未遂政变”。Automattic 旗下包括 Wordpress.com、Tumblr 和 Beeper 等业务。本周早些时候 Mullenweg 通过公司 Slack 频道指责首席财务官 Mark Davies 与董事会串通,董事会投票决定由 Davies 担任临时 CEO。两天后,Mullenweg 称董事会已重新达成一致,他本人恢复了对 Automattic 控制。而 Davies 的 Slack 账户则被停用了,Automattic Slack 频道中的所有管理员也都被移除了。

北京全面限制无人机

北京市政府公布了新修订的《北京市无人驾驶航空器管理规定》,全面限制无人机。《规定》将自 2026 年 11 月 15 日起实施。《规定》明确,本市行政区域全域为无人驾驶航空器管制空域,禁止在本市行政区域内实施无人驾驶航空器飞行活动,禁止在本市行政区域内持有、存放无人驾驶航空器及其核心部件,禁止运输、携带无人驾驶航空器及其核心部件进入本市行政区域。《规定》还要求特殊保障单位应当建立安全管理制度,明确管理责任,防止发生安全事件,特殊保障情形的飞行活动严格按照国家有关规定执行。

暴雪宣布 FPS 版《星际争霸》

暴雪宣布了 FPS 版《星际争霸》,游戏仍然处于早期开发阶段,目标发售时间是在 2030 年。暴雪称,新作是一款开放世界、剧情驱动的科幻射击游戏,故事背景设定在《星际争霸 II》事件发生后数十年,是《星际争霸》宇宙中的一款全新作品。RTS 版《星际争霸》于 1998 年发布,2015 年发布了《星际争霸II》三部曲中的第三部《虚空之遗》,时隔 11 年之后宣布的正统续作不再属于 RTS。FPS 版《星际争霸》游戏设定在 Koprulu 星区(Koprulu Sector),时间位于《虚空之遗》剧情结束后的 70 年。

墨西哥毒贩涉足加密货币挖矿

墨西哥贩毒集团涉足了加密货币挖矿业务。墨西哥警方在 Puebla 州的 Sierra Norte 地区发现了一个用电量远超周边村庄的矿场,查获了 300 个 GPU、80 个中压终端设备以及 8 个卫星天线。虽然就国际商业规模而言,该矿场的规模相当有限,但这已是自去年年初以来该地区发现的第四个加密货币矿场。根据区块链分析公司 Chainalysis 对流向非法钱包地址的交易量进行的分析,全球范围内非法加密货币交易在 2025 年增长一倍以上,与犯罪活动相关的地址接收的资金总额达到 1540 亿美元,而前一年仅为 590 亿美元。拉美的贩毒集团也愈来愈频繁利用加密货币转账和挖矿洗钱。电力是加密货币挖矿的最主要成本,挖一枚比特币的成本接近 4.5 万美元。按当前约 7.8 万美元的市场价出售,矿场仍能有可观的利润。如果矿场还能偷电,那么利润会更高。

Waymo 举报了两名携带幽灵枪的青少年乘客

Waymo 举报了两名携带幽灵枪的年轻乘客。事件发生在 9 月 3 日凌晨 4 点前,地点是旧金山的 Richmond 区。Waymo 发言人称,它在检测到乘客携带枪支之后,停下了无人出租车,通知了执法部门。旧金山警方拘留了两名未成年青少年,一名男孩和一名女孩,搜查汽车后发现了一支已上膛的 AR 风格突击步枪。两名乘客已被送往少年拘留中心。这不是 Waymo 第一次举报乘客,它在今年 7 月曾举报了玩玩具枪的两名青少年乘客。

克雷数学研究所就 Navier-Stokes 问题发表公开声明

OpenAI 本周早些时候宣布通过动用约 1 万个 AI 智能体进行长达 88 小时的攻坚,于 9 月 5 日发现了一个 Navier-Stokes 方程失效的特例。Navier-Stokes 问题是克雷数学研究所列出的七大千禧年数学问题之一,它为每道题的解决提供了一百万美元奖金。目前七大问题只有庞加莱猜想确认解决,但解决该问题的俄罗斯数学家格里戈里·佩雷尔曼拒绝接受该奖。克雷数学研究所就此公开声明,表示将根据其评奖流程确认成果。根据克雷数学研究所的规则,OpenAI 首先需要在同行评审的期刊上发表解题结果,需要至少发表两年并且获得数学界的普遍认可。OpenAI 至今没有给出 Navier-Stokes 问题的证明,因此确认该问题被解决至少需要到 2029 年。

25 名菲尔茨奖得主发表公开信批评 AI 公司

包括陶哲轩、新晋得主邓煜在内的 25 名菲尔茨奖得主发表公开信《A Severe Misalignment of AI in Mathematics》,批评 AI 公司最近的所作所为。公开信称,“过去几个月 LLM 的数学能力有飞跃式提升,甚至达到了能解决数学领域重大悬而未决问题的地步。但各大 AI 公司仅仅将解决数学问题作为基准测试(Benchmark)推动的技术竞赛,却对数学这门科学以及整个数学界构成了伤害。AI 公司的目标与数学界的本质目标之间存在着严重错位(Misalignment)。”这是 AI 影响科学与创意行业乃至整个社会的对齐(alignment)危机的一个缩影。“最近几个月 AI 在解决重大数学难题上取得的突破甚至冲出数学界,登上了大众媒体的头条。然而解决问题仅仅是达成概念理解与深刻洞察这一核心目标的工具和替代指标。在 AI 浪潮中忽略这一点,无异于异化了工具,使其走向核心目标的对立面。事实上,以越来越快的节奏批量生产‘真/假’断言,非但无法为新思想注入生命力,反而可能毁掉孕育创新的沃土。AI 的解答往往发布得过于仓促,甚至没有留出足够时间去编写一份严谨且规范的论文,无法去提炼其中蕴含的新方法与新思想,也无法合理引用前人的相关工作。正如在所有创意行业中发生的一样,这引发了严重的归属权认定与学术剽窃问题。此外,如果没有心怀热忱的数学家去负责对其进行后续开发并融入数学规范体系,AI 所孕育的思想就永远无法真正获得生命,数学家之间至关重要的人际传递纽带也将断裂。”

因 NASA 削减预算 ESA 将独立完成金星探索项目

因特朗普政府削减了 NASA 预算,难以兑现提供合成孔径雷达的承诺,欧洲 ESA 将独立推进金星探索项目 Envision。Envision 轨道探测器任务旨在对金星表面进行测绘,由于金星表面被厚厚的硫酸云层笼罩,需要使用雷达穿透云层。NASA 与 ESA 于 2024 年签署了一份谅解备忘录,NASA 提供美制合成孔径雷达,通过其深空网络提供跟踪与通信支持。作为交换,ESA 将把美国研究人员纳入其团队。然而 2026 年和 2027 年的 NASA 预算被特朗普政府提议大幅削减,虽然国会否决了大部分预算削减方案,但 NASA 的科学预算仍然减少了数亿美元。NASA 的国际合作任务是主要削减对象。

09

APP STORE RANK

09.00
APP STORE RANK
Loading…
TEXT VIEW · TODAY'S DIGEST · 0 HEADLINES ACROSS 8 SOURCES

Hacker News(0)

No items yet for today.

GitHub Trending(0)

No items yet for today.

Product Hunt(0)

No items yet for today.

Hugging Face(0)

No items yet for today.

Techmeme(0)

No items yet for today.

Solidot(0)

No items yet for today.

Startup Archive(0)

No items yet for today.

App Store Rankings(0)

No items yet for today.