#deepmind
4 articles
所有標籤 claude-fable-5 claude-mythos-5 anthropic claude model-release ai-models grok-imagine-video grok-imagine-video-1-5 xai image-to-video ai-video seedance-2 wan-2-7 reve-2-0 reve gpt-image-2 nano-banana-2 nano-banana-pro ai-image image-generation vidu vidu-q3 video-api enterprise b2b nvidia cosmos3 cosmos3-nano world-model physical-ai robotics gemini-omni seedance-2-0 kling-3-0 video-generation model-comparison flux-2 imagen-4 image-api kling-omni storyboarding runway model-marketplace multi-model audio-video technical-breakdown agnes-video agnes-ai pricing leaderboard claude-sonnet-4-8 leak gemini-3-5-flash google google-io agent-tools deepmind gemini-3-5 gemini-3-5-pro gemini-omni-flash gemini-4 claude-mythos gpt-5-5 seedance bytedance gpt-5-6 openai chatgpt alignment gemini omni veo hidream open-source diffusion-transformer multimodal tutorial rumor llm api comparison guide best alternative openrouter aws azure google-cloud inference image-to-3d 3d 3d-generation tripo3d h3.1 pbr quad-mesh announcement wavespeedai multiview-to-3d text-to-3d text-to-image nucleus nucleus-image image-to-image materials texture patina image-to-map game-dev unreal unity blender material-extract video-to-video video-inpainting object-removal masking void sam3 text-to-audio music-generation music-cover style-transfer minimax pixverse animation reference-to-video character-consistency text-to-video transition baidu ernie multilingual chinese fast audio-to-video music-video lip-sync runway-ml image-editing ideogram sora portrait-effect photo-styling parkour action-video talking-photo travel-photo ad-generation spokesperson virtual-try-on fashion ai-avatar free-tools avatar-generator talking-head image-generator video-generator kling predictions deevid-ai wan alibaba video-editing video-edit wan-2.7 video-extend veo3 start-end-to-video kuaishou elements vace video-joiner wavespeed-ai gemma-4 on-device-ai audio-converter audio-processing file-conversion image-converter image-processing face-blur privacy video-converter video-processing ai-tools 4k midjourney flux nano-banana seedream best-ai-image-generator pixverse-v6 audio video-effects glm zhipu-ai gpt deepseek ai-news phota image-enhance upscaler image-quality photorealistic camera-control vfx cybersecurity ai-music suno lyria magihuman davinci sand-ai digital-human davinci-magihuman professional ai-image-generator qwen-image pollo-ai lovart freepik ai-video-generator best-ai-video-generator higgsfield kling-image-o3 ai-image-generation girl-filter face-transformation portrait smile-filter photo-editing watermark-removal sora-alternative sora-shutdown pika grok ltx veo-4 photo-colorizer colorize photo-restoration body-swap face-swap portrait-transfer prismaudio video-to-audio foley ai-audio sound-generation hunyuan audio-generation v2a iclr recraft recraft-v4 text-to-vector svg design dall-e people-remover inpainting fotor photo-editor content-creation desktop-app mp3 wav flac aac png jpg webp heic mp4 mov avi webm janitor-ai media-io video-editor m2.7 ai-model agent coding benchmark age-filter entertainment aging dog-selfie pet-content gender-swap ghibli-filter anime studio-ghibli midjourney-v8 stable-diffusion best-tools ai-content-detector content-moderation content-safety nsfw-detection text-moderation image-moderation video-moderation moderation-api developer-guide sketch-to-video infinitetalk celebrity-look-alike face-recognition clothes-changer fat-filter meme fortune-teller math-solver education story-generator creative-writing review baseten 2026 canva fal-ai fireworks-ai leonardo-ai modal gpu-cloud replicate cloudflare runpod together-ai ai-research helios bitdance bitdance-14b autoregressive qwen-image-2 typography skyreels skyreels-v3 talking-avatar portrait-animation soulx flashhead soulx-flashhead real-time streaming sora-2 ai-images q3 wavespeed-desktop playground batch-processing lora android mobile workflow ai-pipeline ffmpeg audio-conversion image-conversion video-conversion video-merge video-trimming video-enhancement video-upscale inworld tts text-to-speech voice-ai coming-soon gpt-image kimi moonshot-ai ai-assistant local-ai personal-ai prediction genie-3 interactive-environments mova clawdbot personal-assistant automation chatbot javascript typescript sdk python speculation ai-collaboration productivity ai-agents no-code app-builder development apple background-remover face-enhancer image-enhancement image-eraser inpaint tools claude-code codex ai-coding cursor developer-tools image-enhancer ai-platforms hedra avatars heygen creative video-marketing ideas adobe firefly quality rankings image-upscaling enhancement image-translation localization video-extension video-upscaling developer clipdrop stability-ai dalle deepai performance black-forest-labs vertex-ai infrastructure hailuo-ai hugging-face tencent imagen text-rendering kling-ai luma-ai dream-machine serverless nightcafe ai-art pika-labs lm-arena digital-twins tips video-production synthesia dalle-3 prompting avatar multi-modal aimlapi byteplus comfyui dreamina kie-ai openart poyo-ai skywork topaz upscaling qwen training fine-tuning depth controlnet pose upscale outpaint canny lightricks sdxl background-removal marketing event e-commerce product-photography mochi cogvideo social-media instagram
Gemini 3.5 Flash 正式發布——Flash 級模型在 Agent 基準測試上超越 Pro 級
Gemini 3.5 Flash 於 I/O 2026 正式推出,預設開啟思考模式,每百萬 token 定價 $1.50/$9,在 MCP Atlas 及大多數 Agent 測試套件上超越 Claude Opus 4.7 與 GPT-5.5。本文分析 Flash 的領先之處、落後之處,以及如何部署。
4 分鐘閱讀
Gemini 3.5 Pro 下個月即將到來——Flash 版本已透露的訊息
Google 在 I/O 2026 上發布了 Gemini 3.5 Flash,並將 Pro 版本推遲至六月。Flash 已在程式碼和代理基準測試上超越 Gemini 3.1 Pro,但在複雜推理方面有所退步——這正是 Pro 需要彌補的差距。以下是已知資訊、未知資訊,以及如何提前規劃。
3 分鐘閱讀
Google I/O 2026 的 Gemini 4.0:哪些已確認、哪些來自匿名消息、開發者真正需要關注什麼
Google I/O 今日上午 10 點(太平洋時間)正式開幕。關於新版 Gemini 的賽前報導從「漸進式 3.5 更新」到「深度整合的完整 Gemini 4.0」眾說紛紜。以下整理哪些是官方確認的資訊、哪些來自匿名消息來源,以及模型卡片發布後開發者應立即評估的七個面向。
2 分鐘閱讀
Google DeepMind Genie 3:創造互動環境的世界模型
探索Google DeepMind的Genie 3,這是一個突破性的世界模型,能從文字提示實時生成互動、可導航的虛擬環境。
1 分鐘閱讀