New Models
AI Minute Newsroom
2026-08-22
An eight-billion-parameter model that reads pictures and draws them, at 4K, with an Apache licence — and the community had it running in two days
SenseTime put the weights for SenseNova-U1.5-8B-MoT on Hugging Face on 19 August under an Apache-2.0 licence, followed by a fine-tuned variant the same day and a set of LoRAs on 20 August. It is a single native multimodal model rather than a pipeline: the same eight billion parameters that look at an image also generate and edit one, an approach the lab calls NEO-unify. The release notes claim six specific gains over the previous version — better composition and material rendering, more legible Chinese and English text in posters and infographics, more efficient native 4K generation, edits that leave the untouched parts of a picture alone, better adherence to instructions that stack up object counts and spatial relationships, and finer control through bounding boxes and reference images. What is easier to check is what happened next. Within roughly forty-eight hours the community had shipped GGUF and FP8 conversions, an ONNX export and a ComfyUI package, so the model runs on ordinary desktop setups. The main repository is at about 950 downloads and 91 likes; reference code is on GitHub and the weights are mirrored on ModelScope. One honest limit: SenseTime published its benchmark comparisons as charts rather than tables, and no independent evaluation has been run yet, so treat the ranking claims as the lab's own.
Why it mattersThe interesting part is not the scoreboard, it is the shape. Most image tools you can download are one model that draws and another that looks; this is one set of weights doing both, small enough to run on a machine you already own, under a licence that lets you sell what you make with it. Native 4K matters for the same unglamorous reason: it is the difference between a picture you can use for a poster and one you have to upscale. And the two-day gap between a Chinese lab posting weights and someone in a Discord running them in ComfyUI is now the normal speed of this ecosystem — a release is no longer an announcement, it is a download that lands on strangers' machines the same week.
✓ Verified · 4 sources
▶ Related video: New #1 Open Source Image AI? | SenseNova-U1 Mac & Windows Guide & TESTS
Read in the app — free, in 9 languages
Related stories
Google's open models have been downloaded a billion times, and outsiders have built 100,000 versions of them
2026-08-22DeepSeek's cheap workhorse can now see — and on agent tasks that need eyes it says it is close to Anthropic's best
2026-08-21Show the robot once — three to twelve seconds — and it gets the job right 59 times out of 100 with no training at all
2026-08-21Six months ago Hollywood sent ByteDance a cease-and-desist. This week they signed a truce — and no money changed hands.
2026-08-20The model invents its own tasks, builds the rig to test them, then trains on the results — and DeepReinforce gave the weights away
2026-08-20