TL;DR
이번 기간 X 포스트에서는 Alibaba가 Qwen 시리즈(예: Qwen3.8 2.4T, Qwen-Image-3.0, Qwen-Audio-3.0-TTS)와 Zvec 0.6.0을 중심으로 'Token Plan'으로 모델·미디어 통합 결제 체계를 제시했으며, AntLingAGI는 124B MoE 기반의 Ling-3.0-flash(활성화 5.1B, 256K 컨텍스트)를 여러 플랫폼에서 무료 제공해 에이전트형·장문 처리 워크로드 지원을 확장했다. Tencent는 AI-Infra-Guard와 CubeSandbox v0.6.0 공개로 인프라 가시성·샌드박스 운영을 개선했고, Kimi Code CLI는 다중 모델 바인딩·환경변수 설정 등 개발자 편의성을 보강했다. 이 흐름은 대형 모델의 운영·결제 단순화, 장문·에이전트 숙련도 향상, 인프라 보안·배포 편의성 강화라는 실무적 전환을 가속화하지만 실제 비용·성능 영향은 플랫폼별 세부 과금·실행환경에 따라 달라질 수 있다.
𝕏 실시간 트렌드 토픽
📈 Alibaba의 Token Plan·Qwen3.8·미디어 모델 업데이트포스트 6
Alibaba가 'Token Plan'을 내세워 Qwen, Wan, HappyHorse, DeepSeek, GLM 등 텍스트·이미지·비디오·오디오 모델 접근을 단일 크레딧 풀로 통합했다. Ali_TongyiLab은 Qwen3.8(2.4T) 오픈 웨이트 공개와 Qwen-Image-3.0, Qwen-Audio-3.0-TTS, Zvec 0.6.0(검색·배포 개선) 출시를 알렸다. Token Plan은 월 $4부터 시작한다고 안내해 소규모 팀의 초기 비용 진입장벽을 낮추려는 목적을 드러냈다.
- Token Plan은 Qwen 계열 및 타사 모델을 하나의 크레딧 풀로 묶어 모델별 구독을 대체하는 청구 방식을 제시했다.
- Ali_TongyiLab은 Qwen3.8(2.4T) 오픈 웨이트와 Qwen-Image-3.0, Qwen-Audio-3.0-TTS 출시를 공지해 멀티모달 역량을 확장했다.
- Zvec 0.6.0은 검색 속도와 배포 단순화를 목표로 'faster retrieval and simpler deployment' 업데이트를 포함했다.
원문 트윗 2개 보기
Alibaba Cloud
@alibaba_cloud
AI tools multiply fast. Bills multiply faster. The Alibaba Cloud Token Plan replaces that sprawl with one credit pool across Qwen, Wan, HappyHorse, DeepSeek, and GLM. Track what you ship, not what you're subscribed to. Plans start from $4 USD for the first month. Get started today: https:// click.alibabacloud.com/m/20000000854/ #AlibabaCloud #TokenPlan #AI #Qwen #HappyHorse #Wan #DeepSeek #GLM #CloudAI
Alibaba Cloud
@alibaba_cloud
AI tools multiply fast. Bills multiply faster. The Alibaba Cloud Token Plan replaces that sprawl with one credit pool across Qwen, Wan, HappyHorse, DeepSeek and GLM. Track what you ship, not what you're subscribed to. Get started today: https:// click.alibabacloud.com/m/20000000854/ #AlibabaCloud #TokenPlan #AI #Qwen #HappyHorse #Wan #CloudAI
📈 Ling-3.0-flash 공개: 124B MoE·256K 컨텍스트로 에이전트 워크로드 지원포스트 5
AntLingAGI와 협업 생태계에서 Ling-3.0-flash가 공개되어 여러 플랫폼(OpenRouter, Nous Portal, Vercel 등)에서 무료 제공 기간을 운영했다. 모델은 124B 파라미터 MoE 구조며 토큰당 활성화되는 파라미터는 약 5.1B로 표기되었고, 256K 컨텍스트 지원으로 장문·에이전트 작업의 실무 적용성을 강조했다. 무료 제공은 일정 기간에 한정되어 배포·평가 기회를 제공했다.
- 모델 사양으로 124B 총 파라미터, 활성화 파라미터 약 5.1B가 공개되었고 MoE 설계가 토큰 효율을 목표로 한다고 발표됐다.
- 256K 컨텍스트 지원이 명시되어 장문 대화·에이전트 상태 보존에서 장점이 부각됐다.
- OpenRouter·Nous·Vercel 등 여러 플랫폼에 초기 무료 제공을 배포해 개발자 접근성을 높였다.
원문 트윗 2개 보기
Ant Ling
@AntLingAGI
A lovely video it is Now let's build something fun with Ling-3.0-flash on Hermes and keep all the good "memories" Thanks @NousResearch for the day0 support!
Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters and 5.1B active, it's quick to run and built for agent workloads: coding, search, research, and tool use. Try it free today at http:// portal.nousresearch.com/signup

Ant Ling
@AntLingAGI
Extremely honored to work with Vercel Gateway, a partner with great taste! Thanks @vercel_dev and @novita_labs for the Day0 launch~
Free through 8/3: Ling 3.0 Flash from @AntLingAGI is on AI Gateway. A MOE model for agentic workflows, with support for thinking and non-thinking modes. 𝚖𝚘𝚍𝚎𝚕: '𝚒𝚗𝚌𝚕𝚞𝚜𝚒𝚘𝚗𝚊𝚒/𝚕𝚒𝚗𝚐-𝟹.𝟶-𝚏𝚕𝚊𝚜𝚑-𝚏𝚛𝚎𝚎' https:// vercel.com/changelog/ling -3-0-flash-is-now-available-on-ai-gateway …
📈 Tencent의 오픈소스 인프라 툴: AI-Infra-Guard·CubeSandbox v0.6.0포스트 2
Tencent는 AI-Infra-Guard라는 자체 호스팅 레드팀 플랫폼을 공개해 인프라 지문·CVE 스캔, MCP 서버 및 에이전트 스킬 감사, LLM·에이전트 워크플로 레드팀을 로컬 Docker 및 Web UI/API로 수행할 수 있게 했다. 또한 CubeSandbox v0.6.0은 Kubernetes 통합과 볼륨 지원을 추가해 샌드박스를 일반 K8s 워크로드처럼 배포·운영하게 했고, 인용된 지표로는 Sub-60ms 콜드스타트와 메모리 오버헤드 5MB 미만이 제시됐다.
- AI-Infra-Guard는 인프라 취약성 스캔·MCP 서버 감사·에이전트 워크플로 레드팀 기능을 로컬에서 수행할 수 있는 오픈소스 툴로 공개됐다.
- CubeSandbox v0.6.0은 Kubernetes 네이티브 배포와 외부 스토리지 연결을 위한 볼륨 프레임워크를 도입했다.
- 공개된 성능 지표(인용)는 Sub-60ms 콜드스타트와 약 5MB 미만 메모리 오버헤드로 경량 샌드박스 운영을 목표로 한다.
원문 트윗 2개 보기
Tencent AI
@TencentAI_News
AI-Infra-Guard is an open-source, self-hosted red-teaming platform for AI builders. • Scan AI infrastructure for fingerprints and CVEs • Audit MCP servers and Agent Skills • Red-team Agent workflows and LLMs • Run locally with Docker; integrate via Web UI or API Build with more visibility: https:// github.com/Tencent/AI-Inf ra-Guard …
Tencent AI
@TencentAI_News
Open source update CubeSandbox v0.6.0 is live with Kubernetes and Volume support, led by these two community requests: Cube no longer needs a separate sandbox cluster. It can be deployed, scheduled and operated like a regular K8s workload. Bring your own storage to Cube. The E2B-compatible Volume framework uses Create, Destroy, Attach and Detach hooks to decouple sandbox storage from any specific backend. : https:// github.com/TencentCloud/C ubeSandbox …
We just open-sourced Cube Sandbox! An instant, concurrent, secure and lightweight sandbox runtime for AI Agents. Built with RustVMM and KVM, it achieves the perfect balance of security and performance: → Sub-60ms cold start (2.5-50x faster) → Under 5MB memory overhead per
➖ 개발자 도구 업데이트: Kimi Code CLI 개선 및 호환성 패치포스트 2
Kimi Code CLI 0.29.1이 릴리스되어 전역 기본 MCP 서버 타임아웃 설정, OAuth 없이 웹 검색·웹 페치 서비스 구성용 환경변수 추가, 신규 서브에이전트용 2차 모델 바인딩(실험적) 등을 도입했다. 관련 문서에서는 OpenAI 호환 엔드포인트에서 추론 리즈닝 필드 이름 차이를 보완하는 버그 픽스도 공개됐다.
- CLI 0.29.1은 전역 default MCP 서버 타임아웃을 config.toml과 환경변수로 설정할 수 있게 했다.
- 웹 검색·웹 페치 서비스를 OAuth 없이 환경변수로 구성 가능하도록 변경해 자동화 환경에 유리하다.
- 실험적 기능으로 새로 생성되는 서브에이전트에 대한 secondary-model 바인딩과 에이전트 전용 모델 오버라이드를 지원한다.
원문 트윗 2개 보기

Kimi Developers
@KimiDevs
Kimi Code CLI 0.29.1 Features Add global default MCP server timeouts in config.toml and env vars. Add environment variables to configure the web search and web fetch services without OAuth login. Add experimental secondary-model bindings for newly spawned subagents, including per-agent model preferences and subagent-only model overrides.

Kimi Developers
@KimiDevs
Bug Fixes Fix loss of thinking content with OpenAI-compatible endpoints that return reasoning under a different field name (e.g. newer vLLM). See more https:// moonshotai.github.io/kimi-code/en/c onfiguration/config-files.html#secondary-model …
AI 요약 · 북마크 · 개인 피드 설정 — 무료
출처 · 인용 안내
인용 시 "요약 출처: AI Trends (aitrends.kr)"를 표기하고, 사실 확인은 원문 보기 기준으로 진행해 주세요. 자세한 기준은 운영 정책을 참고해 주세요.