update readme

Files changed (1) hide show

README.md CHANGED Viewed

@@ -39,7 +39,7 @@ model-index:
 ## Model Description
-**ClipTagger-12b** is a 12-billion parameter vision-language model (VLM) designed for video understanding at massive scale. Developed by [Inference.net](https://inference.net) in collaboration with [Grass](https://grass.io), this model was created to meet the demanding requirements of trillion-scale video frame captioning workloads
 **ClipTagger-12b exceeds or matches the performance of GPT-4.1 and Claude 4 Sonnet, while costing 15x less per generation.**
@@ -234,4 +234,4 @@ Contact us at [[email protected]](mailto:[email protected]) for a free
 ## License
-This model is released under the Apache-2.0 license, allowing for commercial use and modification with proper attribution.

 ## Model Description
+**ClipTagger-12b** is a 12-billion parameter vision-language model (VLM) designed for video understanding at massive scale. Developed by [Inference.net](https://inference.net) in collaboration with [Grass](https://grass.io), this model was created to meet the demanding requirements of trillion-scale video frame captioning workloads.
 **ClipTagger-12b exceeds or matches the performance of GPT-4.1 and Claude 4 Sonnet, while costing 15x less per generation.**
 ## License
+This model is released under the Apache-2.0 license, allowing for commercial use and modification with proper attribution.