GyaanSetu AI

AI, machine learning and LLM insights.

1515 articlesDeep, practical knowledge

Anthropic એ ઓક્ટોબર 2025 થી ભારતીય આવકમાં બમણો વધારો કર્યો અને બેંગલુરુમાં ઓફિસ ખોલી

ભારત ક્લોડ (Claude) માટેનું બીજા નંબરનું બજાર બનતા Anthropic એ બેંગલુરુમાં નવી ઓફિસ ખોલી રહી છે. એશિયા પેસિફિક ક્ષેત્રમાં આ તેમની બીજી ઓફિસ છે. તેમની પ્રથમ...

AI · 1 min read

Nemotron 3.5 Lightning હવે AWS પર ઉપલબ્ધ, એન્ટરપ્રાઇઝ માટે LLM હાર્ડવેર ખર્ચમાં મોટો ઘટાડો કરશે

ડેવલપર્સ હવે માત્ર એક ક્લિક સાથે SageMaker JumpStart કન્સોલ પરથી સીધું Nemotron 3.5 Lightning શરૂ કરી શકે છે, જેનાથી અલગ GPU ક્લસ્ટર્સ, ડ્રાઇવર ઇન્સ્ટોલેશન અથવા કસ્ટમ કન્ટેનર્સની જરૂરિયાત દૂર થશે. તેની કિંમત ટોકન-આધારિત છે અને પ્રદેશ મુજબ અલગ-અલગ હોઈ શકે છે, તેથી ટીમોએ પ્રોડક્શનમાં ઉપયોગ કરતા પહેલા લેટન્સી અને ચોકસાઈની ચકાસણી કરી લેવી જોઈએ.

AI · 2 min read

AI Context Compression Keeps Only 17% of User Rules, Study Finds

Penn State researchers found that when LLMs compress conversation history, they discard session constraints like “never use my name” or “confirm before changes,” leaving just 17% of such rules intact and jeopardizing security.

AI · 3 min read

180M-Parameter LLM Runs on a $10 ESP32-P4 Microcontroller

The p-for-llm project packs a 180.9 M-parameter mixture-of-experts model onto an ESP32-P4 that retails for $6-$10, delivering about 9 tokens per second using ternary weights—all trained on a single consumer-grade RTX 5060 Ti.

AI · 2 min read

57.8% of AI Agent Skills Violate Specs, Audit Finds 2,465 Listings

The audit of 2,465 public AI-agent skills revealed 57.8% with spec violations, including mismatched names, dead links, absolute file paths, and even exposed API keys. Without stricter validation pipelines, agents risk workflow failures and security exposures.

AI · 4 min read

DeepSeek V4 Pro GA Cuts Reasoning Tokens 18-62% – Lower Bills for Developers

The GA release, announced without a changelog, slashes reasoning-token usage by up to 62%, turning into immediate cost savings for metered users, but it also drops the model’s built-in refusal behavior and requires the “thinking” flag to be off for correct JSON output.

AI · 2 min read

Google Bolsters Chrome with Ex-Relay CEO to Push AI Agents into the Browser

Relay will cut off free users on Aug 15 and paying customers on Sep 14, ending its AI-automation service. Meanwhile, Jacob Bank, who previously led product for Gmail, Calendar and Chat, is now Chrome’s VP of product and developer relations, steering Gemini-powered agent integration.

AI · 5 min read

Flock Makes Case-Number Entry Mandatory After 46 Officer Abuse Reports

The watchdog changes come after a Washington Post investigation uncovered 46 instances of officers using Flock’s 120,000-camera network for personal stalkings and other unauthorized searches, prompting the company to force a case-number field that currently isn’t validated.

AI · 3 min read

એપલ અલીબાબા સાથે ચીનમાં પોતાનું પ્રોપ્રાઈટરી AI મોડલ રજિસ્ટર કરનાર પ્રથમ યુએસ કંપની બની

ગયા મહિને એપલે સત્તાવાર રીતે તેની ઓન-ડિવાઇસ જનરેટિવ AI સેવા ચીની નિયમનકારો પાસે રજિસ્ટર કરી હતી, જે એક યુએસ ટેક જાયન્ટ માટે ઐતિહાસિક પાલન માઈલસ્ટોન છે. આ નવું મોડલ આગામી iOS અપડેટમાં “Apple Intelligence” ફીચર્સને સજ્જ કરશે.

AI · 5 min read

Kog Claims 30x Faster LLM Decoding on Existing Nvidia H200 GPUs

Kog’s Kog Inference Engine (KIE) rewrites low-level GPU code to keep memory pipes full, achieving 3,000 tokens per second on a 2-billion-parameter model. Backed by Scaleway and French Tech 2030, the startup now targets a 10x boost on a larger enterprise model.

AI · 5 min read