Connect with us

Artificial Intelligence

iOS 27 Siri: The 8 Upgrades That Could Finally Make Apple’s Assistant Smart

Published

on

iOS 27 Siri: The 8 Upgrades That Could Finally Make Apple’s Assistant Smart

For years, Siri has been the subject of jokes in the world of AI assistants. Despite being a pioneer, it has consistently lagged behind rivals like Google Gemini and Anthropic’s Claude. However, the narrative might be on the verge of a dramatic shift. The upcoming iOS 27 Siri overhaul, expected at WWDC 2026, promises to be the most significant in the assistant’s history. This article compiles the most compelling rumors about what could make this upgrade a blockbuster.

1. A Dedicated Siri Chatbot App

One of the most anticipated changes is the arrival of a standalone Siri application. Currently, Siri lacks a persistent chat interface, memory, and conversation history. Building on this, the new app is rumored to resemble popular chatbots, offering a text-and-voice interface with saved conversations and pinned chats. This move would not only modernize the user experience but also allow Apple to update Siri’s intelligence independently of full operating system releases, enabling much faster iteration in the competitive AI landscape.

2. A Surprising Brain Transplant: Google Gemini

Perhaps the most surprising rumor points to a deep technical partnership. Reports suggest Apple may power the next generation of its AI models with Google Gemini technology. This collaboration makes strategic sense. Google’s Gemini Nano model excels at on-device processing, aligning perfectly with Apple’s privacy-focused, local computation preferences. Therefore, this partnership could provide the foundational intelligence Siri has critically needed.

Why This Partnership Makes Sense

Apple’s preference for on-device AI to protect user privacy is well-known. Google’s advancements in efficient, powerful models that run directly on a device offer a path for Siri to become smarter without compromising Apple’s core principles. This could be the key to closing the capability gap.

3. True Personal Context Understanding

This feature has been promised and delayed, but iOS 27 might be its moment. The new iOS 27 Siri is expected to finally understand your personal context. Imagine asking, “What was that restaurant my partner texted me about yesterday?” and receiving a precise answer from your messages, not a generic web search. This means Siri could pull data from your emails, calendar, photos, and files to execute tasks and answer complex, multi-step questions. This level of integration is what transforms an assistant from a novelty into a daily necessity.

4. Revolutionary On-Screen Awareness

Another long-awaited upgrade is sophisticated on-screen awareness. While Siri currently has limited visual intelligence, the new system could understand content on your display contextually. For instance, you could look at a text with an event and say, “Add this to my calendar for tomorrow,” and Siri would execute the command seamlessly. This extends to saving passes to Wallet, reading nutrition labels, or adding contacts. It’s a subtle feature that, once experienced, becomes indispensable. For more on how AI integrates with daily tasks, see our guide on Apple Intelligence features.

5. Seamless Cross-App Functionality

This could be the most transformative upgrade. Siri may gain the ability to perform actions within and across applications without you opening them. The secret is the App Intents framework, which allows developers to expose core app functions to Siri. Consequently, you could voice-command Siri to edit a photo in one app and share it via another, or draft and send an email—all through a single request. This deep, system-level access is a potential advantage over cloud-based chatbots like ChatGPT.

6. A New Home in the Dynamic Island

According to reliable sources like Mark Gurman, Siri might get a visual redesign centered on the Dynamic Island. Instead of taking over the entire screen, Siri interactions could live in this compact area, showing progress for longer requests. Additionally, system-wide “Ask Siri” and “Write with Siri” buttons are rumored. This design philosophy makes the assistant ever-present yet unobtrusive.

7. Integration of Third-Party AI Models

Apple appears to be embracing an open ecosystem. The Extensions system, which currently works with ChatGPT, might expand to include other models like Claude, Google Gemini, and Perplexity. You could theoretically route specific queries to your preferred AI from within Siri. However, it’s important to note these third-party tools likely won’t have the same deep system access as the native iOS 27 Siri, functioning more as specialized knowledge partners.

8. Natural Language Shortcuts Creation

The powerful Shortcuts app has often been intimidating for average users. iOS 27 could democratize automation by letting Siri build shortcuts via natural language. Simply say, “Create a shortcut that texts my ETA home when I leave the office,” and Siri would construct it. This approach could unlock the app’s potential for millions, moving automation from a niche tool to a mainstream convenience. Discover more automation tips in our article on iPhone productivity hacks.

Can Apple Finally Deliver?

Given past delays on Siri improvements, skepticism is warranted. Promises made for earlier iOS versions were pushed back. Nevertheless, the scale of rumors this time—encompassing a new app, a major partnership, and system-wide integration—suggests a coordinated relaunch, not a minor update. The reported features address Siri’s core weaknesses: intelligence, context, and accessibility.

Ultimately, WWDC 2026 will reveal if this is the year Siri sheds its outdated skin. If even half these features materialize, the iOS 27 Siri upgrade could redefine our relationship with the iPhone, turning a passive tool into an active, intelligent partner. The wait for a truly smart assistant may finally be over.

Continue Reading
Click to comment

Leave a Reply

Your email address will not be published. Required fields are marked *

Artificial Intelligence

Chinese open-weight models are cheap. Washington is deciding what that costs.

Published

on

Chinese open-weight models

Kimi K3 lands, and the policy debate reignites

On July 16, Moonshot AI dropped Kimi K3, the largest open-weight model ever released. Within days, it had reopened a policy argument in Washington that had been dormant for a year. The question for enterprises evaluating Chinese open-weight models this month isn’t about benchmarks. It’s about whether using one will still be straightforward a year from now.

The outcome will affect procurement decisions well outside the United States. The mechanisms under discussion — federal procurement rules, export blacklists, security advisories — travel through the same cloud providers that serve most of the world.

The immediate trigger: a post by Dean W. Ball, OpenAI’s head of strategic futures and until recently a senior AI adviser in the Trump White House.

Ball’s forecast: regulatory risk, not a ban

Ball’s assessment of the model was largely positive. He called it a very good model whose performance he didn’t think could be explained away by distillation. He also noted it seemed ‘very token hungry’ and wasn’t obviously cheap to run — a useful caution, given K3 launches with maximum reasoning effort as its only setting and bills output at $15 per million tokens.

Then he predicted the Trump administration would eventually decide its best strategy was to create regulatory risk around Chinese open-weight models. Not a ban, which he called one of the dumber motifs in AI policy, but soft guidance from agencies suggesting such models may contain backdoors. ‘It needn’t be that well justified,’ he wrote. Enough uncertainty, and regulated enterprises retreat on their own.

Why Chinese open-weight models are a commercial problem first

The reaction was fierce, and it came from Americans rather than Beijing. David Sacks, co-chair of the President’s Council of Advisors on Science and Technology, said he couldn’t tell whether Ball was confessing to a regulatory capture strategy or predicting one. Either way, weaponising regulatory uncertainty as a competitive tool should be unacceptable, Sacks argued.

He added that the leading closed labs, already a duopoly in model revenue, want the government to remove their open-source competition. Yann LeCun and Martin Casado argued that open and proprietary development can coexist. Ball later clarified he had been forecasting rather than recommending, and walked back the claim that open weights necessarily slow the field down.

Underneath the personalities is an arithmetic problem. Closed labs need revenue per token to justify the capital they are raising for data centres. Cheaper open-weight models compress that revenue without reducing how much AI gets used — the point Snorkel AI co-founder Braden Hancock put to TechCrunch. The routing data already shows the shift: open-weight models handled 29% of tokens through Vercel’s production gateway in June, up from roughly a ninth in April, while accounting for under 4% of spending.

That pressure is arriving from inside the American stack. GitHub made Moonshot’s Kimi K2.7 Code generally available in the Copilot model picker on July 1, hosted on Microsoft Azure. The Information reports Microsoft is now adding K3 to Azure and evaluating whether it can run Copilot features currently handled by OpenAI and Anthropic models, with potential inference savings of up to $600 million.

Microsoft has confirmed neither the figure nor which features. It’s an evaluation, not a deployment. But it’s the largest customer of both American frontier labs, pricing the alternative.

The security argument, taken seriously

Commercial motive does not make the security concern fake. The strongest version of it deserves stating. Open weights cannot be recalled. Once a model is downloaded and running inside thousands of organisations, no vendor can patch it, revoke it, or push a fix — a materially different risk profile from a hosted API. Model behaviour is harder to audit than model code: a fine-tune can carry biases or failure modes that no licence inspection would reveal.

NIST has previously found security vulnerabilities in DeepSeek’s open models. For regulated industries, questions about training data provenance and content handling are live regardless of where a model was built.

The counterargument is about proportionality rather than dismissal. Georgetown research fellow Sam Bresnick has argued that halting Nvidia H200 sales to China would slow Beijing considerably more than banning open models Americans want to use — targeting the input rather than the output. Ball himself conceded a version of this in his second observation, attributing China’s open-weight strategy partly to a lack of domestic compute for serving customers. That would make it an unintended byproduct of US export controls in the first place.

What is actually likely to happen

Axios reported on July 20, citing people close to the administration, that Commerce last year weighed adding Chinese AI labs to the Entity List. The NSA and the Office of the National Cyber Director considered issuing an advisory on Chinese AI lab threats. The White House considered an executive order making US companies liable for breaches if they used Chinese models. Officials concerned about stifling innovation killed all of it.

With adviser Sriram Krishnan gone and security hawks louder, the effort has revived. But the described approach is procurement rules, Entity List threats and public pressure rather than prohibition. ‘What’s actually happening is slower and more durable,’ one source told Axios. Neither the White House nor Commerce responded to Axios’s requests for comment. Politico reports Commerce will not move imminently.

Impact on buyers outside the US

For buyers outside the US, the exposure is indirect but real. A rule written for American regulated industries and federal procurement does not bind a Malaysian bank or an Indonesian telco. The hyperscalers are the transmission line.

Most enterprises in this region reach Kimi K3 through Azure, AWS or Google Cloud rather than Moonshot’s own API. If Washington makes hosting Chinese open-weight models uncomfortable enough for those providers, the model quietly leaves the catalogue in Kuala Lumpur at the same time it leaves it in Virginia.

Ball anticipated this in his own post, noting that regulators would not want to push so hard that hyperscalers stop serving Chinese models altogether. That would only drive startups toward less reputable providers. The obvious hedge is to hold your own copy. Moonshot publishes K3’s weights on July 27, and from that point the model cannot be withdrawn from anyone who has downloaded it.

But as covered previously, K3 is a difficult model to self-host. Moonshot recommends serving it across 64 or more accelerators, and the weights alone come to roughly 1.4TB. For most companies, the fallback is theoretical.

That leaves a narrower question than the headlines imply. Not whether Chinese open-weight models are safe or permitted. But whether the specific model you build on will still be in your cloud provider’s catalogue in twelve months, and what it would cost you to move if it isn’t. That’s a due-diligence question, and it’s answerable today.

For more on the technical side, see our analysis of the Kimi K3 open-weight model and its memory-focused architecture. Also explore how Washington AI policy is shaping global tech procurement.

Continue Reading

Artificial Intelligence

Why the Open Source AI Boom Isn’t Squeezing Anthropic — at Least Not Yet

Published

on

open source AI Anthropic

The Two-Speed AI Economy Nobody’s Talking About

Here’s a riddle for the AI era: If companies are ditching expensive frontier models for cheaper open source alternatives, why is Anthropic still raking in more than half of all AI spending on major platforms?

That contradiction sits at the heart of a provocative new argument from Decagon CEO Jesse Zhang. In a post titled “Everyone is wrong about open source AI in the enterprise,” Zhang proposes that frontier labs and open source models aren’t really competing. They’re playing different roles in a single lifecycle.

Expensive frontier models handle the messy, high-risk early stages of a new use case. Once the process is proven and predictable, companies hand it off to leaner, cheaper open source models. The result? Frontier spending barely dips, because new discovery projects keep popping up to replace the ones that mature.

“The frontier labs will keep owning discovery,” Zhang writes. “Open source will increasingly own production.”

What the Data Actually Shows

Zhang doesn’t offer hard numbers, but the data is easy to find — and it largely backs up his thesis.

Take Vercel‘s AI gateway dashboard. Over the past week, DeepSeek has surged to the lead in token volume, processing just over a third of all tokens flowing through Vercel’s infrastructure. Z.ai, the lab behind the popular GLM-5.2 model, jumped to fourth place in the same period.

But scroll down to spend, and the picture flips. Anthropic still accounts for more than half of all AI spending on the platform. That share has slipped slightly — partly because Anthropic raised prices — but hasn’t collapsed.

OpenRouter tells a similar story across a broader, slightly less enterprise-focused slice of the market. DeepSeek V4 Flash dominates by raw usage, processing 5.3 trillion tokens weekly. The most popular frontier model, Opus 4.8, handles just over 2 trillion. But the price gap is enormous: Opus costs roughly 23 times more per token ($1.37 per million tokens versus DeepSeek’s 6 cents). That means Opus likely still captures the majority of actual dollars spent.

And that’s before factoring in Nvidia’s Nemotron, which is poised to leapfrog competitors thanks to Nvidia’s deep enterprise relationships and the model’s extreme adaptability.

Why Frontier Labs Aren’t Panicking

The numbers don’t fully prove Zhang’s lifecycle theory, but they do explain why Anthropic isn’t sweating the open source surge — at least not yet.

One reason: the total pool of AI-addressable problems is expanding so rapidly that frontier labs can maintain their position simply by dominating new, unproven use cases. Every time a mature workflow migrates to open source, a fresh batch of harder problems appears to take its place.

Another explanation: some use cases are genuinely too difficult for lighter models. Even as clients experiment with cheaper alternatives, they keep a foot in the frontier door for the toughest tasks. That creates a sticky, high-margin revenue base that open source models can’t easily erode.

What This Means for Enterprise AI Buyers

For companies building on AI, the implication is clear: don’t treat frontier and open source models as an either/or choice. Use frontier models to explore and validate. Once the process is stable, switch to open source for production. It’s a hybrid strategy, not a migration.

This two-tiered economy could become a stable feature of the AI market. Frontier labs keep the premium pricing they need to fund R&D. Open source models get the volume that drives ecosystem growth. And enterprises get a cost-effective path from experimentation to deployment.

As recently as last September, many analysts — including this one — predicted that foundation labs would end up as commodity providers, selling “coffee beans to Starbucks” while the application layer captured the value. Some of that prediction came true: vertical AI startups did switch to lighter models, and the economics of “GPT wrapper” companies have remained stable.

But we’re also seeing that frontier providers have held onto the most desirable part of the marketplace: the premium token price. And that doesn’t look likely to change anytime soon.

For a deeper look at how companies are balancing cost and capability, check out our analysis of enterprise AI adoption strategies. And for more on the specific models driving this shift, see our breakdown of how DeepSeek is reshaping the open source AI landscape.

Continue Reading

Artificial Intelligence

Your AI Year in Review: Anthropic Launches Claude Reflect, a Usage Dashboard With a Wellness Twist

Published

on

Claude Reflect

Anthropic just gave Claude users a new kind of year-in-review tool — and it’s designed to make you think twice about how much you lean on AI.

It’s called Claude Reflect. Think of it as your personal AI usage analytics dashboard, but with a twist: instead of just showing you numbers and topics, it pushes you to reconsider your relationship with the technology. Available now in beta, the feature rolls out to free, Pro, and Max users who have enabled memory on Anthropic’s Anthropic platform.

Reflect isn’t called “Claude Wrapped,” even though it does the same seasonal recap that streaming services and AI tools have made famous. The name matters. Anthropic wants this to be more than a vanity metric dump. It’s a prompt for mindfulness.

What Claude Reflect actually shows you

Head to Settings in the Claude web or desktop app and you’ll find Reflect waiting. It generates a summary of your activity over one, three, six, or twelve months. The breakdown goes beyond counting chats.

It surfaces the topics you engage with most, spots usage patterns, and categorizes your interactions using Anthropic’s 4D AI Fluency Framework. Those four dimensions are: delegation, description, discernment, and diligence. So instead of just seeing “you talked about coding 40% of the time,” you get a structured view of how you work with the AI — whether you’re handing off tasks, asking for explanations, or critically evaluating outputs.

Think of it as a report card for your AI habits, not just a spreadsheet of queries.

Privacy and exclusions

Anthropic is careful about what gets counted. Incognito conversations and any health-integration chats are excluded entirely. The company also states that the data stays inside your dashboard and is not used for any other purpose. That’s a meaningful distinction at a time when every AI company is hungry for training data.

The feature was developed in collaboration with MIT Media Lab, the Digital Wellness Lab at Boston Children’s Hospital, and the Family Online Safety Institute. That lineup signals that the wellness angle isn’t an afterthought — it’s baked into the design from the start.

The wellness angle: Why it stands out

Here’s the part that makes you stop. An AI company building a tool that actively nudges you to use AI less? That’s rare. And honestly, it’s refreshing.

Reflect surfaces questions like: “What’s one thing you want to keep doing yourself, even if Claude could do it faster?” It lets you set quiet hours, and it can schedule nudges that remind you to step away from the screen. A time-spent view is coming soon, which will track how many minutes or hours you spend inside Claude conversations.

All of this feels counterintuitive for a business that makes money when people use its product more. But it also aligns with a growing conversation around digital wellness — the idea that technology should serve us, not consume us. Anthropic is betting that users will appreciate a tool that respects their autonomy, even if it means slightly less engagement.

How to get started with Claude Reflect

If you’re already a Claude user with memory enabled, you can access Reflect right now through the Settings menu in the web or desktop apps. It’s in beta, so expect some rough edges and iterative updates. The time-spent view isn’t live yet, but it’s on the roadmap.

For new users, enabling memory is the first step. Once that’s on, Reflect will begin tracking your patterns and building your personalized dashboard. The summaries are generated on-device or in your account — Anthropic says the data doesn’t leave your dashboard.

What this means for the AI industry

Anthropic’s move is a small but telling signal. Most AI companies are racing to increase usage metrics — more conversations, longer sessions, higher retention. Reflect flips that script by asking: are you using AI well, not just a lot?

It’s too early to tell whether users will embrace the nudge to take breaks or ignore it. But the feature itself is a bet that trust and transparency matter more than raw engagement numbers. In an industry that’s been criticized for addictive design patterns, that’s a notable stance.

If you’re curious about how your own AI habits stack up, or just want a tool that occasionally tells you to log off, Claude Reflect is worth a look. It’s one of the few dashboards that might actually make you feel better about your screen time — not worse.

Continue Reading

Trending