<!-- mobian-agent-page publisher="dailydev" canonical="https://daily.dev/posts/neweyes-is-trying-to-remove-the-prompt-from-multimodal-ai-instead-of-opening-an-app-describing-a--nfojukrlg" -->

---
title: NewEyes is trying to remove the prompt from multimodal...
description: Discussion about &quot;NewEyes is trying to remove the prompt from multimodal AI.

Instead of opening an app, describing a task, and attaching an image, the camera...
canonical: https://daily.dev/posts/neweyes-is-trying-to-remove-the-prompt-from-multimodal-ai-instead-of-opening-an-app-describing-a--nfojukrlg
twitter:card: summary_large_image
twitter:site: @dailydotdev
og:type: website
og:site_name: daily.dev
og:title: NewEyes is trying to remove the prompt from multimodal AI.

Instead of opening an app, describing a task, and attaching an image, the camera itself becomes the trigger. 

Visual perception identifies a possible need, persistent memory adds context, and the system decides whether anything should happen next. 

Which shows up as shopping what you see, styling what you wear, redesigning your room, counting your meals, and flagging a thirsty plant before you ask.

This is a meaningful interface shift, because prompts currently tell the model both what matters and what the user permits it to do.

A proactive system has to infer the first while remaining conservative about the second.

Once the balance is achieved, the camera moves from a feature to the everyday layer, where personal AI is actually there. | daily.dev
og:description: Discussion about &quot;NewEyes is trying to remove the prompt from multimodal AI.

Instead of opening an app, describing a task, and attaching an image, the camera...
og:url: https://daily.dev/posts/neweyes-is-trying-to-remove-the-prompt-from-multimodal-ai-instead-of-opening-an-app-describing-a--nfojukrlg
og:image: https://api.daily.dev/og/posts/nFOjUKrLG.png
og:image:alt: NewEyes is trying to remove the prompt from multimodal AI.

Instead of opening an app, describing a task, and attaching an image, the camera itself becomes the trigger. 

Visual perception identifies a possible need, persistent memory adds context, and the system decides whether anything should happen next. 

Which shows up as shopping what you see, styling what you wear, redesigning your room, counting your meals, and flagging a thirsty plant before you ask.

This is a meaningful interface shift, because prompts currently tell the model both what matters and what the user permits it to do.

A proactive system has to infer the first while remaining conservative about the second.

Once the balance is achieved, the camera moves from a feature to the everyday layer, where personal AI is actually there.
og:image:width: 1200
og:image:height: 630
og:locale: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# NewEyes is trying to remove the prompt from multimodal AI.

Instead of opening an app, describing a task, and attaching an image, the camera itself becomes the trigger. 

Visual perception identifies a possible need, persistent memory adds context, and the system decides whether anything should happen next. 

Which shows up as shopping what you see, styling what you wear, redesigning your room, counting your meals, and flagging a thirsty plant before you ask.

This is a meaningful interface shift, because prompts currently tell the model both what matters and what the user permits it to do.

A proactive system has to infer the first while remaining conservative about the second.

Once the balance is achieved, the camera moves from a feature to the everyday layer, where personal AI is actually there.

**[Rohan Paul](https://daily.dev/sources/rohanpaul_ai)** · 0 upvotes · 0 comments

## Full article

daily.dev links to this article rather than hosting it. Read it at the original source: <https://x.com/rohanpaul_ai/status/2084679429037633898>

## Similar posts on daily.dev

- [Don’t just attend KubeCon \+ CloudNativeCon, Merge Forward your experience\!](https://daily.dev/posts/don-t-just-attend-kubecon-cloudnativecon-merge-forward-your-experience--l0rpp73x8) · CNCF · 1 upvotes · 0 comments
- [Announcing H2 2026 KCDs](https://daily.dev/posts/announcing-h2-2026-kcds-m96goajm1) · CNCF · 1 upvotes · 0 comments
- [Two months of Open Community Groups](https://daily.dev/posts/two-months-of-open-community-groups-asf52zhbs) · CNCF · 0 upvotes · 0 comments
- [CNCF Unveils Schedule for KubeCon \+ CloudNativeCon Europe 2026](https://daily.dev/posts/cncf-unveils-schedule-for-kubecon-cloudnativecon-europe-2026-ikhcoa5cb) · CNCF · 2 upvotes · 0 comments
- [CNCF Debuts KubeCon \+ CloudNativeCon Japan 2026 Schedule](https://daily.dev/posts/cncf-debuts-kubecon-cloudnativecon-japan-2026-schedule-xp5pyudub) · CNCF · 1 upvotes · 0 comments

---

[View this post on daily.dev](https://daily.dev/posts/neweyes-is-trying-to-remove-the-prompt-from-multimodal-ai-instead-of-opening-an-app-describing-a--nfojukrlg)

```json
{"@context":"https://schema.org","@graph":[{"@type":"Organization","@id":"https://daily.dev/#organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180},"sameAs":["https://twitter.com/dailydotdev","https://github.com/dailydotdev","https://www.linkedin.com/company/daily-dev-ltd"]},{"@type":"WebSite","@id":"https://daily.dev/#website","url":"https://daily.dev","name":"daily.dev","publisher":{"@id":"https://daily.dev/#organization"},"potentialAction":{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https://daily.dev/search?q={search_term_string}"},"query-input":"required name=search_term_string"}}]}
{"@context":"https://schema.org","@type":"TechArticle","headline":"NewEyes is trying to remove the prompt from multimodal AI.\n\nInstead of opening an app, describing a task, and attaching an image, the camera itself becomes the trigger. \n\nVisual perception identifies a possible need, persistent memory adds context, and the system decides whether anything should happen next. \n\nWhich shows up as shopping what you see, styling what you wear, redesigning your room, counting your meals, and flagging a thirsty plant before you ask.\n\nThis is a meaningful interface shift, because prompts currently tell the model both what matters and what the user permits it to do.\n\nA proactive system has to infer the first while remaining conservative about the second.\n\nOnce the balance is achieved, the camera moves from a feature to the everyday layer, where personal AI is actually there.","url":"https://daily.dev/posts/neweyes-is-trying-to-remove-the-prompt-from-multimodal-ai-instead-of-opening-an-app-describing-a--nfojukrlg","mainEntityOfPage":{"@type":"WebPage","@id":"https://daily.dev/posts/neweyes-is-trying-to-remove-the-prompt-from-multimodal-ai-instead-of-opening-an-app-describing-a--nfojukrlg"},"datePublished":"2026-08-04T16:34:49.288Z","dateModified":"2026-08-05T07:39:53.699Z","description":"Discussion about \"NewEyes is trying to remove the prompt from multimodal AI.\n\nInstead of opening an app, describing a task, and attaching an image, the camera...","image":"https://media.daily.dev/image/upload/s--VDukGCjf--/f_auto/v1722860399/public/Placeholder%2002","thumbnailUrl":"https://media.daily.dev/image/upload/s--VDukGCjf--/f_auto/v1722860399/public/Placeholder%2002","isAccessibleForFree":true,"articleSection":"Rohan Paul","inLanguage":"en","publisher":{"@type":"Organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180}},"author":{"@type":"Organization","name":"Rohan Paul","logo":"https://media.daily.dev/image/upload/s--elRVZ626--/f_auto,q_auto/v1783603589/logos/rohanpaul_ai?_a=BAMAMicg0","url":"https://daily.dev/sources/rohanpaul_ai"},"commentCount":0,"discussionUrl":"https://daily.dev/posts/neweyes-is-trying-to-remove-the-prompt-from-multimodal-ai-instead-of-opening-an-app-describing-a--nfojukrlg","interactionStatistic":[{"@type":"InteractionCounter","interactionType":{"@type":"LikeAction"},"userInteractionCount":0},{"@type":"InteractionCounter","interactionType":{"@type":"CommentAction"},"userInteractionCount":0}],"keywords":""}
{"@context":"https://schema.org","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https://daily.dev"},{"@type":"ListItem","position":2,"name":"Rohan Paul","item":"https://daily.dev/sources/rohanpaul_ai"},{"@type":"ListItem","position":3,"name":"NewEyes is trying to remove the prompt from multimodal AI.\n\nInstead of opening an app, describing a task, and attaching an image, the camera itself becomes the trigger. \n\nVisual perception identifies a possible need, persistent memory adds context, and the system decides whether anything should happen next. \n\nWhich shows up as shopping what you see, styling what you wear, redesigning your room, counting your meals, and flagging a thirsty plant before you ask.\n\nThis is a meaningful interface shift, because prompts currently tell the model both what matters and what the user permits it to do.\n\nA proactive system has to infer the first while remaining conservative about the second.\n\nOnce the balance is achieved, the camera moves from a feature to the everyday layer, where personal AI is actually there."}]}
```

