
David Pogue: '125 Tests of the New AI Siri'
David Pogue conducted 125 tests on the new AI-powered Siri to evaluate its accuracy and capabilities. While the assistant failed in a few instances, it successfully completed many tasks that surprised the reviewer. Pogue encourages users to overcome the habit of manually opening apps and learn to trust the system to save time.
- ▪David Pogue performed 125 specific tests on the final version of the new AI Siri.
- ▪The test cases included tasks confirmed by Apple, suggestions from Reddit users, and Pogue's own experiments.
- ▪Siri failed a handful of times but also accomplished surprising feats that impressed the tester.
- ▪Pogue advises users to practice using Siri to build trust and reduce the time spent fumbling with apps.
Hacker News (AI / LLM) files mainly under ai. We currently carry 5,888 of its stories.
Story provenance
Source · retrieval · rights · ranking — open for full record
inspect →
Story provenance
Attribution is not the same as permission. This drawer separates discovery metadata, excerpts, WeSearch-generated summaries, reuse status, and whether the publisher receives the visit. Nothing here claims a legal grant the publisher has not made.
Record
| Original publisher | Daring Fireball |
| Canonical URL | https://daringfireball.net/linked/2026/09/21/pogue-siri-ai-125 |
| Publication time | Mon, 21 Sep 2026 17:56:52 +0000 |
| Retrieval time | 2026-09-21T18:08:49.156Z |
| Last seen | 2026-09-21T18:08:49.156Z |
| Headline source | Publisher (no WeSearch rewrite) |
| Excerpt source | publisher body |
| Excerpt method | First ~120 words (~800 chars) of extracted publisher body, fair-use limited. |
| Summary | WeSearch · cerebras-chat (WeSearch summarizer) |
| Summary source text | contentText |
| Citation coverage | Summary is a WeSearch-generated derivative; primary citation is the original publisher URL. |
| Cluster | jktqcdlWQZFq · 1 stories |
| Cluster logic | Grouped by semantic title/content similarity across sources within a rolling window. Same-publisher template collisions are excluded from coverage comparison. |
| Ranking reason | Story pages are not engagement-ranked. Hub feeds use recency, with optional source-diversified chronological ordering (cap consecutive stories per source). No personalized ranking. |
| Publisher visit | Yes — open original |
| Substitutes article? | No — link-out required for full text |
Rights status (four layers)
WeSearch handling by dimension
| Indexing | May the item be indexed (stored, ranked, made findable)? | Allowed |
| Snippet | May a short excerpt of the publisher's text be shown? | Allowed |
| AI summary | May WeSearch generate its own short summary of the article? | Limited |
| Retrieval / RAG | May the content be exposed for third-party retrieval-augmented generation? | Not asserted |
| Model training | May the content be used to train AI models? | Not asserted |
| Commercial reuse | May the content be reused commercially? | Not permitted |
Basis: Derived from the published RSS/Atom feed. Contact: [email protected]. Reviewed: 2026-07-24.
Opening excerpt (first ~120 words) tap to expand
The hardest part will be getting out of the habit of opening apps, and remembering that you have Siri. Give it a few tries. Learn to trust it. You’ll save incredible amounts of time and fumbling. I used the beta-test version all summer, growing more and more excited. Now that the final version is out, I wanted to see how often it gets things right. It is AI, after all. So I ran it through 125 tests. Some of these tasks are things Apple said are possible. Some ideas, I picked up from people on Reddit trying stuff. Some, I just wondered if Siri could do. It failed a handful of times. But a few other times, it absolutely blew me away. It did amazing things I bet you never suspected Siri could do.
Excerpt limited to ~120 words for fair-use compliance. The full article is at Daring Fireball.