Creeta — AI developer tools & ecosystem news
NeMo out, GGUF in: how parakeet.cpp ports NVIDIA ASR to C++
parakeet.cpp v0.1.0: NVIDIA Parakeet in GGUF — no NeMo needed. CMake steps, quant tradeoffs, and whisper.cpp status.
Is Omni's conversational video editor as good as the demos?
Gemini Omni in Google Flow: credit costs, regional limits, and iterative editing — no callable API yet.
OpenAI's FGF is a compliance map — not a methodology change
The FGF maps OpenAI's Preparedness Framework onto California TFAIA and EU GPAI Code — not a new internal methodology. Published May 28, 2026.
Windsurf is Devin Desktop now. Cascade has 27 days left.
Windsurf is now Devin Desktop: Agent Command Center, Spaces, and Devin Local replacing Cascade by July 1.
Meta's always-on pendant will record everyone in the room — not just you
An internal Alex Himel memo, reported by The Information, reveals Meta's AI pendant roadmap: ambient audio capture, real-time transcription, and a Wearables for Work subscription tier — built on the Limitless acquisition.
3,000 tok/s on MI300X by deleting the kernel scheduler
Kog AI monokernel: ~3,000 tok/s on AMD MI300X by eliminating kernel launches. Technical read with caveats.
Windsurf is Devin now — Cascade retires July 1
Cognition renamed Windsurf to Devin Desktop June 2. What changed, what broke, and what IT admins need to do now.
Nemotron 3 Ultra went live June 4. Here's the call that works.
NVIDIA Nemotron 3 Ultra GA June 4: how to call via NIM/OpenRouter, hardware floor, and the base-checkpoint caveat.
Composer 2.5 hits near-frontier at 60× lower spend
Composer 2.5: third on the Artificial Analysis Coding Index at $0.07/task vs $4.10 for its nearest rival. Billing choice, effective prompting, and what the independent scores actually show.
4 GitHub stars, voice interviews with Ollama: that's GrillKit
Apache 2.0 interview trainer with Whisper voice input, Ollama or cloud LLM support, and local session history. No SaaS, no registration required.
RDNA3 cuts llama.cpp KV VRAM 47% — and CUDA has no equivalent
RDNA3 bit-packing cuts llama.cpp KV VRAM 47% on RX 7900. Flags, VRAM math, and TurboQuant for 4.9× compression.
NodeCartel is dark. Cross-host AI orchestration: who delivers.
NodeCartel is unreachable. Kore.ai, CrewAI Cloud, Northflank, and AgentNode Pro compared for cross-host AI scheduling.