New 3D editors, open medical AI, AI symphony, Qwen 3.8, Wan Animate 2: AI NEWS

This week’s AI news highlights major advancements including Symphony Gen for orchestral music composition, Alibaba’s One Animate 2 for detailed character animation, and the powerful open-source language model Qwen 3.8 Max rivaling top AI systems. Additionally, breakthroughs span cyclone forecasting, medical imaging analysis, robotics, and innovative frameworks for AI agents, collectively pushing the boundaries of creativity, scientific research, and autonomous AI capabilities.

This week in AI news has been packed with groundbreaking developments across various domains. One highlight is Symphony Gen, an AI capable of composing full orchestral symphonies by first generating a harmony skeleton and then expanding it into detailed arrangements. This model is lightweight and available for local use, offering musicians and composers a powerful tool for music creation. In 3D modeling, the new multi-agent CAD system (MAC) efficiently generates printable 3D CAD files from text prompts, significantly reducing costs and token usage compared to previous models, and is model-agnostic for flexible integration.

Alibaba introduced One Animate 2, an advanced animation system that animates characters from photos using reference videos, including detailed hand, finger, and facial movements. It supports multiple characters, irregular body proportions, and adjustable camera angles, with a lighter version enabling real-time streaming. Vocal Render, another notable AI, generates highly realistic and expressive singing voices from lyrics and MIDI melodies, outperforming competitors and offering open-source models that can be trained on various languages. Tencent’s Hunyan 3D Buffalo stands out as a unified 3D model generator, editor, and segmenter, capable of text-based edits and part separation, with code release anticipated soon.

In the realm of large language models, Alibaba’s Quen 3.8 Max, a 2.44 trillion parameter model, is set to be open-sourced soon and demonstrates performance rivaling top models like GPT-5.6 and Fable 5. It excels in autonomous software engineering tasks, including self-improving code systems and scientific research, marking a significant step toward autonomous AI-driven innovation. Google DeepMind’s Weather Next 2 AI advances cyclone and hurricane forecasting by combining track and intensity predictions into a single model, offering more accurate forecasts up to 15 days ahead with much lower computational requirements, and is fully open-sourced for public use.

OpenAI revealed that an internal model, Astra, solved ten longstanding open mathematical problems across various fields at a remarkably low computational cost, signaling a new era of AI-assisted scientific breakthroughs. Alibaba’s Clinfusion model offers holistic medical understanding by analyzing diverse medical imaging types and generating clinical reports, outperforming many proprietary models and available in two sizes for different hardware capabilities. Robotics also saw exciting progress with Persona AI’s teleoperated humanoid robot performing precise welding tasks, UB Robotics’ swarm intelligence coordinating multiple robots in warehouses, and Xiaomi’s Robotics 1 model trained on extensive human and robot data to handle everyday object manipulation, with models openly released for developers.

Additional innovations include Leap Talk, a real-time talking avatar generator noted for its speed; Higsfield’s Seed Dance 2.5 video generator offering extensive control over multi-shot narratives; Meta’s Musepark 1.2 focused on coding and agentic workflows with a large context window; and the Long Horizon Harness framework that improves AI agent performance on complex, long-duration tasks by dividing roles into manager, executor, and auditor. The experimental Big Bang model explores self-evolving AI by generating its own challenging training data, showing significant performance gains. Overall, these developments showcase rapid advancements in AI capabilities across music, 3D modeling, language understanding, scientific research, medical analysis, robotics, and creative content generation.