luokai

@luok.ai

For more AI&Tech content, check here www.luok.ai 🍎Apple Die Hard Fan| 苹果骨灰粉 🤖GenAI Observer | GenAI观察者 👨🏻‍🎤Cutting Edge Tech Enthusiast | 科技爱好者

The Seedance 2 model is incredibly powerful, completely overshadowing all other models. This is an original video I created in just one day, though the music was previously made using Suno.

Finally, it’s official: Apple’s next AI leap is… built on Google’s Gemini. 🤯 Apple and Google have signed a multi-year agreement: future Apple Foundation Models will be based on Gemini models and Google Cloud technology.

A Meta Quest open-source MR app turns your room into a language lab. Spatial Lingo shows how mixed reality + AI can teach vocab by labeling your real world—now open-source.

Niji 7 just landed. The latest Niji focuses on sharper eyes, tighter coherency, and better prompt adherence. It keeps legacy flags and adds sref tweaks for style control. After 18 months of training, this release targets fewer misses and more faithful outputs for anime creators.

AI video is stuck in “fun”—the next leap is making it truly useful. Over two years since Sora’s splashy debut, generative video is dazzling but still struggles to land in everyday workflows. The missing piece?

HRM^2Avatar turns a single iPhone scan into a high‑fidelity, real‑time digital human—on mobile. ⚡ It combines mesh‑driven clothing deformation with illumination‑aware Gaussians, learning textures from static shots and dynamics from motion clips.

Meshy 6 just raised the ceiling for AI 3D characters—cleaner topology, sharper forms, truer anatomy. Upgrades that you’ll feel in production: — Refined Geometry: Cleaner, anatomically correct meshes that deform better when rigged.

Kling just gave video creators granular control over pacing—start/end frames with selectable durations. Key stats Start/End Frame duration: 3–10s Resolution modes: 1080p and 720p (same features)

Gemini 3’s new Deep Think mode tackles problems by exploring multiple hypotheses at once ⚡ Positioned as Google’s most advanced reasoning tier, Deep Think runs iterative rounds to refine outputs—especially for complex tasks like code visualization, prototyping, and nuanced analysis.

Kling AI just leveled up avatars to full 5‑minute performances. Wild. Avatar 2.0 is upgraded and expressive—built to handle explainers, ads, songs, and stories like real characters. Key stats 5‑minute avatar acts - Max expressions

Kling AI drops VIDEO 2.6: native audio meets coherent visuals — finally, story-first AI video. This isn’t “just a clip.” VIDEO 2.6 outputs synchronized sound, lip‑sync, and consistent scene logic, enabling short‑film narratives.

Kling launches IMAGE O1: input anything, understand everything, generate any vision. 🚀 A full-stack revamp from generation to editing with superb consistency, precise modification, powerful stylization, and max creativity.

A brand‑new creative engine just dropped: Kling O1 can turn any input into vision. With true multimodal understanding, Kling O1 unifies text, images, and videos to speed up creation. Techy, fast, and playful—input anything, it understands everything.

Runway just unveiled a frontier video model: Gen‑4.5—raising the bar for motion, fidelity, and prompt control. Positioned as their new world‑modeling foundation, Gen‑4.5 prioritizes physical accuracy, scene choreography, and efficient pre/post‑training.

PixVerse V5.5 just made text-to-video feel like a film studio in your pocket. 🎬 Key stats Multi-shot in one tap - Dynamic SFX auto-added - Single-prompt flow reduces setup steps dramatically

Vidu AI drops a turbocharged Q2 Image Model—fast, consistent, and 4K-ready. Key stats As fast as 5s generation - 4K - Unlimited image generation for members until Dec 31 - New users: code VIDUQ2RTI for bonus credits

ChatGPT Voice now lives inside the chat—no mode-switch gymnastics needed. 🎙 OpenAI rolled out real-time voice + visuals directly in chat on web and mobile. You can talk, watch answers appear, revisit earlier messages, and see images/maps inline.

Gemini just added interactive images to make hard topics click. Google’s Gemini App now turns static diagrams into tappable, explorable visuals for more dynamic learning. Instead of passive watching, you can drill into layers—cells, circuits, and more—built for academic depth with quick feedback.

Come and see the cinematic language and incredible detail in this generated test! 🐋 The image was created with nano banana Pro, and the video was brought to life by Grok.

NotebookLM just dropped a new way to see your research: Infographics. Key stats Rolling out to 100% of Pro users today; free users in the coming weeks. Mobile support (infographics + slide decks) will follow fast.

Introducing a controllable video leap: Multi-Frames by Dreamina. It lets creators define up to 10 keyframes and generate a coherent one-take clip. Transitions are fully adjustable, with prompts guiding how each scene evolves. Up to ~54s per take - 1–6s transitions - Frame-level prompt control.

Gemini’s new image model just leveled up: Nano Banana Pro. Key stats Text: multilingual, long-paragraph legibility. Inputs: blend up to 14 images; keep resemblance for up to 5 people. Access: Gemini app, Ads, Workspace, API, Vertex AI.

I’ve got to say, Grok Image is my AI game-changer of the year. It’s so impressive I was about to subscribe for the video feature, and the fact that it’s free is just mind-blowing! And with the video extension update on the way? Absolutely incredible!”

Apple is quietly reshaping hardware manufacturing with industrial-scale 3D printing — and it’s not a prototype era anymore. This year, Apple moved the entire production of Watch Ultra 3 and titanium Series 11 cases to laser‑printed, recycled aerospace‑grade titanium.