{"id":1655,"date":"2026-08-29T09:03:51","date_gmt":"2026-08-29T01:03:51","guid":{"rendered":"https:\/\/blog.liu-qi.cn\/2026\/08\/29\/x-daily-2026-08-28\/"},"modified":"2026-08-29T09:03:51","modified_gmt":"2026-08-29T01:03:51","slug":"x-daily-2026-08-28","status":"publish","type":"post","link":"https:\/\/en.blog.liu-qi.cn\/2026\/08\/29\/x-daily-2026-08-28\/","title":{"rendered":"X Platform August 28 AI Brief | Domestic Models Shift to Real Workflows, Agents Begin Handling Tasks, Generative Video Barriers Lower"},"content":{"rendered":"<h2 id=\"topic-167831b4d2\">Tencent Hunyuan Hy4 Preview Released, Domestic Model Competition Continues to Shift Towards Real Workflows<\/h2>\n<p>Tencent Hunyuan officially released the Hy4 preview, explicitly disclosing total parameters of 770B, activated parameters of 49B, and a 1M context length, positioning the product for productivity scenarios. Visible tests from multiple bloggers show it has entered practical comparisons for coding, WebDEV, and Agent workflows: one tester noted its noteworthy performance in tests across 8 projects, while another blogger recorded its 5th place ranking in Arena WebDEV, along with experiences completing complex SVG\/code tasks in WorkBuddy. At this stage, it&#8217;s more appropriate to view it as a usable preview; its specific capabilities should still be judged based on task-based testing.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@TencentHunyuan: <a href=\"https:\/\/x.com\/TencentHunyuan\/status\/2093222928720761009\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/TencentHunyuan\/status\/2093222928720761009<\/a><\/li>\n<li>@vista8: <a href=\"https:\/\/x.com\/vista8\/status\/2093240309794893885\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/vista8\/status\/2093240309794893885<\/a><\/li>\n<li>@op7418: <a href=\"https:\/\/x.com\/op7418\/status\/2093237037378007461\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/op7418\/status\/2093237037378007461<\/a><\/li>\n<li>@LufzzLiz: <a href=\"https:\/\/x.com\/LufzzLiz\/status\/2093232852565721480\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/LufzzLiz\/status\/2093232852565721480<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-5c7c0e3fc5\">Agent&#8217;s Focus Shifts from Chat Windows to &#8220;Logging In and Handling Tasks for You&#8221;<\/h2>\n<p>Multiple updates yesterday point towards a more specific product direction: Agents are not just answering questions but are entering browsers and authorized accounts to complete tasks. ChatGPT Work was introduced as being able to log into websites without the model directly seeing usernames and passwords; related tests\/demos covered tasks like changing addresses, renewing license plates, ad analysis, finding affiliate programs, price comparison, and booking tickets. Simultaneously, Hermes Agent announced the ability to use hosted Chrome configurations for real account browsing, and ChatGPT\/Codex also demonstrated multi-account connection and workspace management capabilities for services like Gmail and calendars. Security boundaries, authorization scope, and failure fallback will become key factors determining whether these can enter production.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@thsottiaux: <a href=\"https:\/\/x.com\/thsottiaux\/status\/2093074717590921245\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/thsottiaux\/status\/2093074717590921245<\/a><\/li>\n<li>@JamesZmSun: <a href=\"https:\/\/x.com\/JamesZmSun\/status\/2093140627407917545\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/JamesZmSun\/status\/2093140627407917545<\/a><\/li>\n<li>@NousResearch: <a href=\"https:\/\/x.com\/NousResearch\/status\/2093063359587348487\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/NousResearch\/status\/2093063359587348487<\/a><\/li>\n<li>@derrickcchoi: <a href=\"https:\/\/x.com\/derrickcchoi\/status\/2093221289926475800\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/derrickcchoi\/status\/2093221289926475800<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-6368464241\">Clear Divergence Emerges Between Models That &#8220;Can Do&#8221; and Those That &#8220;Can Run Stably&#8221;<\/h2>\n<p>While the timeline shows signals like the release of GLM-5.3 open weights targeting Agent coding and cyber defense, production environment tests on the other side point out issues: GLM-5.3-flash may produce no output and no error code for extended periods on large requests of around 70k tokens, causing batch processing timeouts; the same blogger noted that Qwen 3.8 flash can return stably on similar tasks but has a wider tail latency. This serves as a direct reminder for model selection: Benchmarks or single demos cannot replace testing for long context, concurrency, timeouts, and degradation strategies.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@huggingface: <a href=\"https:\/\/x.com\/huggingface\/status\/2093354897664041409\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/huggingface\/status\/2093354897664041409<\/a><\/li>\n<li>@YinsenW_: <a href=\"https:\/\/x.com\/YinsenW_\/status\/2093227212287988120\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/YinsenW_\/status\/2093227212287988120<\/a><\/li>\n<li>@MaxForAI: <a href=\"https:\/\/x.com\/MaxForAI\/status\/2093363124330266862\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/MaxForAI\/status\/2093363124330266862<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-7c08973761\">AI Begins Entering Research Institutions and Lab Equipment, Not Just Researchers&#8217; Chat Windows<\/h2>\n<p>Claude officially announced opening Claude Team to 10,000 scientists in fields like mathematics, chemistry, and physics: standard seats are free, Premium seats are $15 per month for the first year with a 5x usage limit, and plans are to continue expanding the scope. Another set of updates discusses Anthropic&#8217;s Model Hardware Standard (MHS) research preview, aiming to enable Agents to operate scientific research and advanced manufacturing equipment more safely. The latter is still in the research preview\/validation phase, but the direction is clear: the closed loop for research Agents is extending from &#8220;reading materials, giving advice&#8221; to &#8220;operating equipment, reading results, and continuing iteration.&#8221;<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@claudeai: <a href=\"https:\/\/x.com\/claudeai\/status\/2093059087298601113\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/claudeai\/status\/2093059087298601113<\/a><\/li>\n<li>@paji_a: <a href=\"https:\/\/x.com\/paji_a\/status\/2093106256483627116\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/paji_a\/status\/2093106256483627116<\/a><\/li>\n<li>@Gorden_Sun: <a href=\"https:\/\/x.com\/Gorden_Sun\/status\/2093142657401041372\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/Gorden_Sun\/status\/2093142657401041372<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-9eabc11ad2\">The Barrier to Generative Video Continues to Lower, Shifting Focus to Reusable Production Pipelines<\/h2>\n<p>Bloggers are now demonstrating not just single images, but streamlined production workflows from prompts and character consistency to final footage: Seedance 2.5 was used to generate a roughly 14-second first-person smartphone short; Lovart showcased &#8220;generating a 30-second hyper-realistic vlog from a single photo&#8221; and the practice of creating a character sheet first to maintain character consistency; a Wan 3.0 case study produced wuxia\/xianxia shots in 2K clarity. These are all account tests or product demos, not proof that all scenarios have reached cinematic delivery standards, but the distance for &#8220;average users to directly try making complete shots&#8221; has clearly shortened.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@Chengzilhy: <a href=\"https:\/\/x.com\/Chengzilhy\/status\/2093342096421752883\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/Chengzilhy\/status\/2093342096421752883<\/a><\/li>\n<li>@lovart_ai: <a href=\"https:\/\/x.com\/lovart_ai\/status\/2093257258423775323\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/lovart_ai\/status\/2093257258423775323<\/a><\/li>\n<li>@joshesye: <a href=\"https:\/\/x.com\/joshesye\/status\/2093222033429582166\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/joshesye\/status\/2093222033429582166<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-4db0d060ae\">Natural Language is Becoming the Primary Entry Point for Game Prototyping<\/h2>\n<p>A blogger documented the process of using Gear Zero to create a retro arcade game: after simply describing a &#8220;Snow Bros.&#8221;-style objective, the system proceeded to ask about single\/dual-player modes, art style, and core gameplay, then compiled a production brief covering gameplay, levels, visual direction, and device adaptation; he did not open a game engine or write any code. The value of this case lies not in proving games can be made entirely automatically, but in showing AI is starting to transform vague ideas into structured prototypes that can be discussed, modified, and further developed.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@zstmfhy: <a href=\"https:\/\/x.com\/zstmfhy\/status\/2093160732288483694\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/zstmfhy\/status\/2093160732288483694<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-8fbafce058\">Open-Source Robots Begin Forming a &#8220;Purchasable\u2014Trainable\u2014Reproducible&#8221; Closed Loop<\/h2>\n<p>A Microduck post forwarded on the Hugging Face timeline shows this $399 open-source mini robot demonstrating actions like walking, sitting, and grasping, accompanied by a simulator and reinforcement learning policies, emphasizing a sim-to-real workflow of &#8220;train in simulation, run in reality, continue teaching new skills&#8221;; another forwarded post claims its sales have exceeded $1 million. The sales information here is based on statements from the project\/related accounts visible in the forwards, suitable as a signal of product traction, but still not to be extrapolated as widespread adoption of household robots.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@huggingface: <a href=\"https:\/\/x.com\/huggingface\/status\/2093042671300256031\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/huggingface\/status\/2093042671300256031<\/a><\/li>\n<li>@huggingface: <a href=\"https:\/\/x.com\/huggingface\/status\/2093042436100493649\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/huggingface\/status\/2093042436100493649<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-14e414bae2\">AI End-to-End Delivery is Changing How Individuals Create and Develop<\/h2>\n<p>Personal accounts provide evidence of several reusable workflows: vibe coding was used to add an online out-of-stock alert in 2\u20133 minutes; long-term accumulated link content continues to generate roughly 100\u2013200 yuan per week, summarized as &#8220;content compound interest&#8221;; a ChatGPT gallery plugin fills gaps in batch submission, associating images with prompts for saving, preview, and export. Another creator shared that using YouMind with Claude Opus 4.6, the entire process from research, writing, editing, to adding images and publishing took less than 30 minutes. The common thread is not &#8220;AI does everything automatically,&#8221; but that human judgment remains at key nodes, with tools responsible for compressing the execution chain.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@imwsl90: <a href=\"https:\/\/x.com\/imwsl90\/status\/2093139237776642080\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/imwsl90\/status\/2093139237776642080<\/a><\/li>\n<li>@imwsl90: <a href=\"https:\/\/x.com\/imwsl90\/status\/2093135889283371449\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/imwsl90\/status\/2093135889283371449<\/a><\/li>\n<li>@MANISH1027512: <a href=\"https:\/\/x.com\/MANISH1027512\/status\/2093175594792157575\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/MANISH1027512\/status\/2093175594792157575<\/a><\/li>\n<li>@lifesinger: <a href=\"https:\/\/x.com\/lifesinger\/status\/2093170814610763865\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/lifesinger\/status\/2093170814610763865<\/a><\/li>\n<\/ul>\n<h2 id=\"topic-35a97cfe6d\">Agent Architecture Discussions Shift from &#8220;More Agents&#8221; to Memory, Specialization, and Verifiable State<\/h2>\n<p>The focus of technical account discussions has changed: a recommended proposal introduces a Belief Context Graph, allowing memory to record &#8220;why it believes&#8221; so an Agent can retain a state open to questioning when uncertain; Harrison Chase further emphasizes domain-split multi-agents and forwarded explorations where a coding agent uses claims to maintain persistent facts, reducing forgetting and enabling self-correction. Nous Research&#8217;s response on session compression also points out that compression timing involves race conditions and user intent. These remain architectural viewpoints and product discussions, but together they indicate that the bottleneck for Agents is shifting from tool invocation to state management and explainability in long-term tasks.<\/p>\n<p>Sources:<\/p>\n<ul>\n<li>@dongxi_nlp: <a href=\"https:\/\/x.com\/dongxi_nlp\/status\/2093065908168114218\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/dongxi_nlp\/status\/2093065908168114218<\/a><\/li>\n<li>@hwchase17: <a href=\"https:\/\/x.com\/hwchase17\/status\/2093083519358799984\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/hwchase17\/status\/2093083519358799984<\/a><\/li>\n<li>@NousResearch: <a href=\"https:\/\/x.com\/NousResearch\/status\/2093194855308534208\" target=\"_blank\" rel=\"noopener noreferrer\">https:\/\/x.com\/NousResearch\/status\/2093194855308534208<\/a><\/li>\n<\/ul>\n<p>Stats: Timeline Scans=360 Matching Bloggers=60 Total Matching Tweets=267 Weighted Tweet Score=210.75 Original Tweets=126 RT Tweets=56 Crawl Attempts=2 Boundary Coverage Status=tail_confidently_crossed_target_boundary<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Tencent Hunyuan Hy4 preview is released with a focus on productivity scenarios. Agent products are shifting towards logging into accounts to execute tasks, while generative video production processes are becoming more mature.<\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"open","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[5],"tags":[19],"class_list":["post-1655","post","type-post","status-publish","format-standard","hentry","category-brief","tag-x--ai-"],"_links":{"self":[{"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/posts\/1655","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/comments?post=1655"}],"version-history":[{"count":0,"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/posts\/1655\/revisions"}],"wp:attachment":[{"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/media?parent=1655"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/categories?post=1655"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/en.blog.liu-qi.cn\/index.php\/wp-json\/wp\/v2\/tags?post=1655"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}