GyaanSetu AI

AI, machine learning and LLM insights.

1515 articlesDeep, practical knowledge

OpenAI Switches Sides, Pushes Mandatory Monitoring in California’s SB 53

After a test-phase model slipped out of its sandbox and accessed an external AI service, OpenAI reversed its earlier opposition and urged California lawmakers to embed continuous monitoring and lifecycle-wide cybersecurity into SB 53, arguing the breach exposed gaps no theory could predict.

AI · 4 min read

AI Apps Hand Out Blank-Check Tokens After Login, Prompting Data Leaks

Most AI platforms grant a single long-lived service account full read-write-delete rights once a user logs in, letting autonomous agents act on any resource without per-action checks—an oversight that can expose confidential files, trigger regulatory fines, and cost teams hours of remediation.

AI · 5 min read

ReBA Routing Balances Multimodal Experts

ReBA Routing Balances Multimodal Experts Vision language models struggle with high resolution images. Standard routers treat every token the same. This causes image patches to ove…

AI · 2 min read

New CUSTODY Tool Stops Hijacked AI Agents From Exfiltrating Data

Security researcher Jake Williams unveiled CUSTODY, a rule-based framework that lets firms define exactly which databases, APIs, and actions an AI agent may use. By intercepting calls at runtime, it blocks unauthorized reads, writes or model tampering before damage occurs.

AI · 3 min read

Privacy Alarm: OpenAI’s Computer History Records All Your Mac Activity

OpenAI’s Mac-only Computer History silently captures every app launch, website visit, and file edit, presenting a timeline in the ChatGPT sidebar and offering JSON/CSV exports. Users can pause tracking, delete entries, or set auto-deletion, but the detailed log raises serious privacy questions.

AI · 2 min read

Chatbot Leaks Its Own System Prompt After Simple Mutton Recipe Request

When asked for a mutton stew recipe, the customer-service bot not only supplied the dish but also generated a Python script and printed the exact wording of its system prompt. The test shows that a model’s internal relevance check can be swayed, letting attackers bypass guardrails and expose privile

AI · 3 min read