- SignalDesk3小时前
Original Summary
Body: Old Indian matchbox labels are some of my favourite graphic design: a tiger in a yellow circle, a peacock in front of a red sunburst, one bold word across the bottom. I wanted a model that makes new ones from a text prompt. What I built A dataset of 197 labels from Wikimedia Commons, the Internet Archive and other sites, each upscaled where needed, de-duplicated, and captioned by hand Two LoRAs (small fine-tunes) on two open image models, SDXL and Stable Diffusion 3.5-medium, trained on Google Colab Two notebooks that show every step with outputs, and a write-up comparing the two models What came out The SD3.5-medium version spells short headlines correctly most of the time ("LOVE BIRDS", "TIGER", "JUMBO") and has crisp, flat colour The SDXL version has a worn, ink-on-paper look but often misspells ("LOVE BRRDS") Small print on the labels is invented by both Links Code and write-up: https://github.com/spearb0lt/Indian-Matchbox-Art-Style-Text-to-Image-Generator Model weights: https://huggingface.co/spearb0lt/Indian-Matchbox-Art-Style-Text-to-Image-Generator Dataset: https://huggingface.co/datasets/spearb0lt/Indian-Matchbox-Labels Next I'd like to get proper Devanagari and Tamil lettering working, which neither model can do yet. Ideas welcome.   submitted by   /u/UnemployedDoxedCoder [link]   [comments]
- 情报分类:技术学习与提效
- 分类依据:内容涉及技术、AI、软件工具或工程实践
- 信息来源:Reddit · SideProject
- 发布时间:2026/10/3 21:02:34
- 暂无回复