<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Ai on myByways</title><link>https://myByways.com/tags/ai/</link><description>Recent content in Ai on myByways</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Sat, 25 Jul 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://myByways.com/tags/ai/index.xml" rel="self" type="application/rss+xml"/><item><title>Migrating from Grav to Hugo</title><link>https://myByways.com/post/migrating-from-grav-to-hugo/</link><pubDate>Sat, 25 Jul 2026 00:00:00 +0000</pubDate><guid>https://myByways.com/post/migrating-from-grav-to-hugo/</guid><description>&lt;p&gt;Due to issues with updating PHP on my shared hosting which I cannot be bothered to resolve, I am now considering migrating from &lt;a href="https://getgrav.org/" rel="external"&gt;Grav&lt;/a&gt; to &lt;a href="https://gohugo.io/" rel="external"&gt;Hugo&lt;/a&gt;.&lt;/p&gt;</description></item><item><title>Chatterbox Turbo TTS with MLX-Audio</title><link>https://myByways.com/post/chatterbox-turbo-tts-with-mlx-audio/</link><pubDate>Tue, 10 Feb 2026 00:00:00 +0000</pubDate><guid>https://myByways.com/post/chatterbox-turbo-tts-with-mlx-audio/</guid><description>&lt;p&gt;In my last post, I tried &lt;a href="https://myByways.com/post/qwen3-tts-with-mlx-audio-on-macos"&gt;Qwen3-TTS&lt;/a&gt;&amp;hellip; This time, I test &lt;a href="https://www.resemble.ai/chatterbox-turbo/" rel="external"&gt;Chatterbox Turbo by Resemble.ai&lt;/a&gt;, which is an open-source, MIT licensed, text-to-speech model with zero-shot cloning.&lt;/p&gt;</description></item><item><title>Qwen3-TTS with MLX-Audio on macOS</title><link>https://myByways.com/post/qwen3-tts-with-mlx-audio-on-macos/</link><pubDate>Mon, 02 Feb 2026 00:00:00 +0000</pubDate><guid>https://myByways.com/post/qwen3-tts-with-mlx-audio-on-macos/</guid><description>&lt;p&gt;Alibaba’s &lt;a href="https://github.com/QwenLM/Qwen3-TTS" rel="external"&gt;Qwen3-TTS&lt;/a&gt; for &lt;strong&gt;Speech Synthesis&lt;/strong&gt; (text-to-speech) was open-sourced (Apache-2.0) on 22 Jan 2026. And within the last couple of weeks, we now have Apple Silicon optimization via &lt;a href="https://github.com/Blaizzy/mlx-audio" rel="external"&gt;MLX-Audio&lt;/a&gt;. Here is code to create an audiobook from an ePub.&lt;/p&gt;</description></item><item><title>Reflections on 2025 through the Word(s) of the Year</title><link>https://myByways.com/post/reflections-on-2025-through-the-words-of-the-year/</link><pubDate>Sun, 04 Jan 2026 00:00:00 +0000</pubDate><guid>https://myByways.com/post/reflections-on-2025-through-the-words-of-the-year/</guid><description>&lt;p&gt;In March 2025, I posted an opinion piece entitled &lt;a href="https://myByways.com/post/my-disillusionment-with-generative-ai"&gt;“My Disillusionment with Generative AI”&lt;/a&gt;. Things have gotten worse since then, to the extent that &lt;strong&gt;“Slop”&lt;/strong&gt; and &lt;strong&gt;“AI Slop”&lt;/strong&gt; have been picked as the Word(s) of the Year for 2025. What follows are my thoughts on the trajectory of technology and, more specifically, AI, as reflected in the Word(s) of the Year (WOTY).&lt;/p&gt;</description></item><item><title>Kokoro TTS and Abogen on Windows</title><link>https://myByways.com/post/kokoro-tts-and-abogen-on-windows/</link><pubDate>Sun, 31 Aug 2025 00:00:00 +0000</pubDate><guid>https://myByways.com/post/kokoro-tts-and-abogen-on-windows/</guid><description>&lt;p&gt;In my last post, I described trying out &lt;a href="https://huggingface.co/hexgrad/Kokoro-82M" rel="external"&gt;Kokoro text-to-speech (TTS) model&lt;/a&gt; via &lt;a href="https://myByways.com/post/running-kokoro-tts-via-macos-containerisation-framework"&gt;Kokoro-FastAPI web UI in a macOS (native) container&lt;/a&gt;. Here, I install &lt;a href="https://github.com/nazdridoy/kokoro-tts" rel="external"&gt;Kokoro-TTS&lt;/a&gt; and &lt;a href="https://github.com/denizsafak/abogen" rel="external"&gt;Abogen&lt;/a&gt; on Windows, to take advantage of my Nvidia GPU.&lt;/p&gt;</description></item><item><title>Running Kokoro TTS via macOS Containerisation Framework</title><link>https://myByways.com/post/running-kokoro-tts-via-macos-containerisation-framework/</link><pubDate>Sun, 17 Aug 2025 00:00:00 +0000</pubDate><guid>https://myByways.com/post/running-kokoro-tts-via-macos-containerisation-framework/</guid><description>&lt;p&gt;A short 2-in-1 post of two things I’ve been meaning to try out on macOS - first, to try the new &lt;a href="https://apple.github.io/container/documentation/" rel="external"&gt;macOS Container framework&lt;/a&gt; on macOS 15.5&amp;hellip; and second, to spin up &lt;a href="https://huggingface.co/hexgrad/Kokoro-82M" rel="external"&gt;Kokoro Text-to-Speech (TTS)&lt;/a&gt; in a container.&lt;/p&gt;</description></item><item><title>My Disillusionment with Generative AI</title><link>https://myByways.com/post/my-disillusionment-with-generative-ai/</link><pubDate>Tue, 18 Mar 2025 00:00:00 +0000</pubDate><guid>https://myByways.com/post/my-disillusionment-with-generative-ai/</guid><description>&lt;p&gt;I’m a techno optimist by default, I do believe technology can solve the problems we face today, and make our lives better. But I do have concerns with the direction we are taking with Generative AI, and the future we are heading towards. This is an opinion piece&amp;hellip;&lt;/p&gt;</description></item><item><title>Using Flux.1 in GGUF format on macOS</title><link>https://myByways.com/post/using-flux1-in-gguf-format-on-macos/</link><pubDate>Tue, 20 Aug 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/using-flux1-in-gguf-format-on-macos/</guid><description>&lt;p&gt;&lt;a href="https://github.com/city96" rel="external"&gt;city96&lt;/a&gt; has published &lt;a href="https://github.com/ggerganov/ggml/blob/master/docs/gguf.md" rel="external"&gt;GGUF&lt;/a&gt; versions of the Flux1.Dev model and T5 XXL text encoder, along with custom nodes to use them in &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; - thought I’d try them on my &lt;strong&gt;M2 Mac mini&lt;/strong&gt;, hoping for faster inference!&lt;/p&gt;</description></item><item><title>Faster Flux.1</title><link>https://myByways.com/post/faster-flux1/</link><pubDate>Wed, 14 Aug 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/faster-flux1/</guid><description>&lt;p&gt;More &lt;a href="https://huggingface.co/black-forest-labs/FLUX.1-dev" rel="external"&gt;Flux.1&lt;/a&gt;-based models! Go faster with FP8 or NF4! New LoRAs and ControlNets! There is quite a bit of interest with this model, as evidenced by the speed of community-led enhancements.&lt;/p&gt;</description></item><item><title>Wow! Flux.1 by Black Forrest Labs</title><link>https://myByways.com/post/wow-flux-by-black-forrest-labs/</link><pubDate>Sun, 04 Aug 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/wow-flux-by-black-forrest-labs/</guid><description>&lt;p&gt;While I have many posts about SDXL, I do not use Stable Diffusion 3 at all - license concerns aside, it is simply not good, and may never get better. But just a few days ago, a new, freely available, offline model that is better than SDXL was released by the team that presented Latent Diffusion and created Stable Diffusion, &lt;a href="https://blackforestlabs.ai/" rel="external"&gt;Flux.1 by Black Forrest Labs&lt;/a&gt;!&lt;/p&gt;</description></item><item><title>Fun with Llama 3 (8B) in Ollama</title><link>https://myByways.com/post/fun-with-llama-3-8b-in-ollama/</link><pubDate>Sun, 21 Apr 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/fun-with-llama-3-8b-in-ollama/</guid><description>&lt;p&gt;I saw a post on Reddit, entitled &lt;a href="https://www.reddit.com/r/LocalLLaMA/comments/1c90gl8/llama_3_rocks_with_taking_on_a_personality/" rel="external"&gt;“Llama 3 rocks with taking on a personality!”&lt;/a&gt;. A fun experiment! I thought to replicate it, blatantly copying the puzzle presented. &lt;a href="https://ai.meta.com/blog/meta-llama-3/" rel="external"&gt;Llama 3, released by Meta&lt;/a&gt; just before the weekend, is impressive.&lt;/p&gt;</description></item><item><title>Testing new PAG and Perp-Neg nodes in ComfyUI</title><link>https://myByways.com/post/testing-new-pag-and-perp-neg-nodes-in-comfyui/</link><pubDate>Wed, 17 Apr 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/testing-new-pag-and-perp-neg-nodes-in-comfyui/</guid><description>&lt;p&gt;I know it’s bad form to start off with a disclaimer: but the truth is, I do not know what I am doing. I am just testing out two new &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; nodes, &lt;kbd&gt;PerturbedAttentionGuidance&lt;/kbd&gt; and &lt;kbd&gt;PerpNegGuider&lt;/kbd&gt;.&lt;/p&gt;</description></item><item><title>Mistral System Message setup to improve Image Generation Prompts</title><link>https://myByways.com/post/mistral-system-message-setup-to-improve-image-generation-prompts/</link><pubDate>Sun, 31 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/mistral-system-message-setup-to-improve-image-generation-prompts/</guid><description>&lt;p&gt;&lt;a href="https://myByways.com/post/llm-prompt-generation-with-ollama-in-comfyui"&gt;In my last post&lt;/a&gt;, I used &lt;a href="https://github.com/if-ai/ComfyUI-IF_AI_tools" rel="external"&gt;ComfyUI-IF_AI_tools&lt;/a&gt; to integrate to the &lt;a href="https://ollama.com/brxce/stable-diffusion-prompt-generator" rel="external"&gt;brxce/stable-diffusion-prompt-generator model&lt;/a&gt; running in &lt;a href="https://ollama.com/" rel="external"&gt;Ollama&lt;/a&gt;. I wonder if I could use the base &lt;a href="https://mistral.ai/news/announcing-mistral-7b/" rel="external"&gt;Mistral 7B model&lt;/a&gt; to help improve my uncreative prompts instead&amp;hellip;&lt;/p&gt;</description></item><item><title>LLM Prompt Generation with Ollama in ComfyUI</title><link>https://myByways.com/post/llm-prompt-generation-with-ollama-in-comfyui/</link><pubDate>Thu, 28 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/llm-prompt-generation-with-ollama-in-comfyui/</guid><description>&lt;p&gt;In my last post, I described running &lt;a href="https://myByways.com/post/a-game-with-mistral-7b-using-ollama"&gt;Mistral, a Large Language Model, locally using Ollama&lt;/a&gt;. To accompany that piece, I created a prompt and manually used AI to generate an image. Today, I’ll wire up a &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; workflow to &lt;a href="https://ollama.com/" rel="external"&gt;Ollama&lt;/a&gt; to do this seamlessly, thanks to &lt;a href="https://github.com/if-ai/ComfyUI-IF_AI_tools" rel="external"&gt;ComfyUI-IF_AI_tools&lt;/a&gt;.&lt;/p&gt;</description></item><item><title>A game with Mistral 7B using Ollama</title><link>https://myByways.com/post/a-game-with-mistral-7b-using-ollama/</link><pubDate>Wed, 20 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/a-game-with-mistral-7b-using-ollama/</guid><description>&lt;p&gt;I keep posting about Stable Diffusion, but I do experiment with &lt;strong&gt;Large Language Models&lt;/strong&gt; too! I do not have much to contribute in this regard, instead, here is the transcript of a game I played with the open source &lt;a href="https://mistral.ai/news/announcing-mistral-7b/" rel="external"&gt;Mistral 7B model&lt;/a&gt; via &lt;a href="https://ollama.com/" rel="external"&gt;Ollama&lt;/a&gt;.&lt;/p&gt;</description></item><item><title>Stable Video Diffusion</title><link>https://myByways.com/post/stable-video-diffusion/</link><pubDate>Sun, 10 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/stable-video-diffusion/</guid><description>&lt;p&gt;More and more AI generated images are shared as short video clips. So, here a quick test of &lt;a href="https://stability.ai/news/stable-video-diffusion-open-ai-video-model" rel="external"&gt;Stable Video Diffusion&lt;/a&gt; - which was released back in November last year. Don’t know why I didn’t post this when I posted about &lt;a href="https://myByways.com/post/generating-animations-videos-with-sdxl-lcm"&gt;AnimateDiff and the Hotshot Motion model&lt;/a&gt; around the same time.&lt;/p&gt;</description></item><item><title>TripoSR image-to-3D-object</title><link>https://myByways.com/post/triposr-image-to-3d-object/</link><pubDate>Sat, 09 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/triposr-image-to-3d-object/</guid><description>&lt;p&gt;Do you want to convert a 2D image into a 3D model &lt;em&gt;auto-magically&lt;/em&gt;? On 5 March 2024, &lt;strong&gt;Stability AI&lt;/strong&gt; and &lt;strong&gt;Tripo AI&lt;/strong&gt; released &lt;a href="https://stability.ai/news/triposr-3d-generation" rel="external"&gt;TripoSR: Fast 3D Object Generation from Single Images&lt;/a&gt; that does exactly that!&lt;/p&gt;</description></item><item><title>Differential Diffusion for in-painting</title><link>https://myByways.com/post/differential-diffusion-for-in-painting/</link><pubDate>Tue, 05 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/differential-diffusion-for-in-painting/</guid><description>&lt;p&gt;&lt;a href="https://differential-diffusion.github.io/" rel="external"&gt;Differential Diffusion&lt;/a&gt; is the newest method (framework) of in-painting without an in-painting model. Instead, all that is needed is a &lt;strong&gt;mask&lt;/strong&gt; (map) where the lighter the area, the greater the re-painting applied.&lt;/p&gt;</description></item><item><title>Generate transparent images with Layer Diffusion</title><link>https://myByways.com/post/generate-transparent-images-with-layer-diffusion/</link><pubDate>Sun, 03 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/generate-transparent-images-with-layer-diffusion/</guid><description>&lt;p&gt;Ever wished you could generate &lt;a href="https://stability.ai/stable-image" rel="external"&gt;Stable Diffusion XL&lt;/a&gt; images with transparent backgrounds? Well, your wish has been answered by the smart people behind the &lt;a href="https://arxiv.org/html/2402.17113v1" rel="external"&gt;Transparent Image Layer Diffusion using Latent Transparency paper&lt;/a&gt;. They have made their code and models available, and what do you know, &lt;a href="https://github.com/huchenlei/ComfyUI-layerdiffusion" rel="external"&gt;Chenlei Hu&lt;/a&gt; has ported it to &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt;!&lt;/p&gt;</description></item><item><title>SDXL-based 4-step models compared</title><link>https://myByways.com/post/sdxl-based-4-step-models-compared/</link><pubDate>Sat, 02 Mar 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/sdxl-based-4-step-models-compared/</guid><description>&lt;p&gt;With the advent of techniques like Adversarial Diffusion Distillation and Latent Consistency models, A.I. image synthesis based on &lt;a href="https://stability.ai/stable-image" rel="external"&gt;Stable Diffusion XL&lt;/a&gt; has been getting faster and faster. Here is just quick comparison of a few models at 4-steps, some of which are fine-tuned and trained for realism.&lt;/p&gt;</description></item><item><title>Consistent portraits revisisted: InstantID</title><link>https://myByways.com/post/consistent-portraits-revisisted-instantid/</link><pubDate>Sun, 25 Feb 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/consistent-portraits-revisisted-instantid/</guid><description>&lt;p&gt;Not long ago, in a attempt to obtain &lt;a href="https://myByways.com/post/consistent-portraits-using-ip-adapters-for-sdxl"&gt;Consistent portraits using IP-Adapters for SDXL&lt;/a&gt;, I &lt;a href="https://myByways.com/post/comparing-face-ip-adapters-for-sdxl"&gt;shared a comparison&lt;/a&gt; between &lt;a href="https://huggingface.co/h94/IP-Adapter" rel="external"&gt;IP-Adapter-Plus-Face&lt;/a&gt; and &lt;a href="https://huggingface.co/h94/IP-Adapter-FaceID" rel="external"&gt;IP-Adapter-FaceID&lt;/a&gt;. Today I’ll look at &lt;a href="https://huggingface.co/InstantX/InstantID" rel="external"&gt;InstantID&lt;/a&gt;.&lt;/p&gt;</description></item><item><title>New Stable Cascase Checkpoints for ComfyUI</title><link>https://myByways.com/post/new-stable-cascase-checkpoints-for-comfyui/</link><pubDate>Sat, 24 Feb 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/new-stable-cascase-checkpoints-for-comfyui/</guid><description>&lt;p&gt;An update to my &lt;a href="https://myByways.com/post/stable-cascade-with-comfyui"&gt;previous post on Stable Cascade with ComfyUI&lt;/a&gt; - instead of requiring four separate model files, we now only need &lt;a href="https://huggingface.co/stabilityai/stable-cascade/tree/main/comfyui_checkpoints" rel="external"&gt;two checkpoints&lt;/a&gt;, and the &lt;a href="https://comfyanonymous.github.io/ComfyUI_examples/stable_cascade/" rel="external"&gt;ComfyUI workflow is now very straightfoward&lt;/a&gt;!&lt;/p&gt;</description></item><item><title>Stable Cascade with ComfyUI</title><link>https://myByways.com/post/stable-cascade-with-comfyui/</link><pubDate>Sun, 18 Feb 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/stable-cascade-with-comfyui/</guid><description>&lt;p&gt;On 12 Feb 2024, &lt;a href="https://stability.ai/" rel="external"&gt;Stability.ai&lt;/a&gt; released &lt;a href="https://stability.ai/news/introducing-stable-cascade" rel="external"&gt;Stable Cascade&lt;/a&gt; “research preview” (non-commercial license), and over the weekend, &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; was updated to support this new model! Time to give it a go!&lt;/p&gt;</description></item><item><title>Comparing face IP-Adapters for SDXL</title><link>https://myByways.com/post/comparing-face-ip-adapters-for-sdxl/</link><pubDate>Sun, 14 Jan 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/comparing-face-ip-adapters-for-sdxl/</guid><description>&lt;p&gt;As a follow up to my last post regarding &lt;a href="https://myByways.com/post/consistent-portraits-using-ip-adapters-for-sdxl"&gt;Consistent portraits using IP-Adapters for SDXL&lt;/a&gt;, this is a short comparison of the two face IP-Adapters for SDXL by &lt;a href="https://huggingface.co/h94" rel="external"&gt;h94 / xiaohu&lt;/a&gt;: namely, &lt;a href="https://huggingface.co/h94/IP-Adapter" rel="external"&gt;ip-adapter-plus-face_sdxl_vit-h.bin&lt;/a&gt; and &lt;a href="https://huggingface.co/h94/IP-Adapter-FaceID" rel="external"&gt;ip-adapter-faceid_sdxl.bin&lt;/a&gt;.&lt;/p&gt;</description></item><item><title>Consistent portraits using IP-Adapters for SDXL</title><link>https://myByways.com/post/consistent-portraits-using-ip-adapters-for-sdxl/</link><pubDate>Sun, 07 Jan 2024 00:00:00 +0000</pubDate><guid>https://myByways.com/post/consistent-portraits-using-ip-adapters-for-sdxl/</guid><description>&lt;p&gt;Getting consistent character portraits generated by SDXL has been a challenge&amp;hellip; until now! &lt;a href="https://github.com/cubiq/ComfyUI_IPAdapter_plus" rel="external"&gt;ComfyUI IPAdapter Plus&lt;/a&gt; (dated 30 Dec 2023) now supports both &lt;a href="https://huggingface.co/h94/IP-Adapter" rel="external"&gt;IP-Adapter&lt;/a&gt; and &lt;a href="https://huggingface.co/h94/IP-Adapter-FaceID" rel="external"&gt;IP-Adapter-FaceID&lt;/a&gt; (released 4 Jan 2024)!&lt;/p&gt;</description></item><item><title>Go even faster with SDXL-Turbo!</title><link>https://myByways.com/post/go-even-faster-with-sdxl-turbo/</link><pubDate>Wed, 29 Nov 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/go-even-faster-with-sdxl-turbo/</guid><description>&lt;p&gt;In the span of a couple of weeks, we got &lt;a href="https://myByways.com/post/crazy-fast-image-generation-with-lcm-lora-for-sdxl"&gt;Crazy fast image generation with LCM LoRA for SDXL&lt;/a&gt;, which led me to ask if I could get &lt;a href="https://myByways.com/post/faster-stable-diffusion-on-mseries-macs"&gt;Faster Stable Diffusion on M-series macs?&lt;/a&gt;. A few hours ago, &lt;strong&gt;Stability.ai&lt;/strong&gt; gave us their response in the form of &lt;a href="https://huggingface.co/stabilityai/sdxl-turbo" rel="external"&gt;SDXL-Turbo&lt;/a&gt;&amp;hellip; and now we go even faster!&lt;/p&gt;</description></item><item><title>Faster Stable Diffusion on M-series macs?</title><link>https://myByways.com/post/faster-stable-diffusion-on-mseries-macs/</link><pubDate>Mon, 27 Nov 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/faster-stable-diffusion-on-mseries-macs/</guid><description>&lt;p&gt;All my recent &lt;a href="https://stability.ai/stable-diffusion" rel="external"&gt;Stable Diffusion XL&lt;/a&gt; experiments have been on my Windows PC instead of my M2 mac, because it has a faster Nvidia 2060 GPU with more memory. But today, I’m curious to see how much faster diffusion has gotten on a M-series mac (M2 specifically).&lt;/p&gt;</description></item><item><title>QR Code Monster for SDXL</title><link>https://myByways.com/post/qr-code-monster-for-sdxl/</link><pubDate>Tue, 14 Nov 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/qr-code-monster-for-sdxl/</guid><description>&lt;p&gt;Time to try another &lt;strong&gt;ControlNet&lt;/strong&gt; for Stable Diffusion XL - &lt;a href="https://huggingface.co/monster-labs/control_v1p_sdxl_qrcode_monster" rel="external"&gt;QR Code Monster v1&lt;/a&gt; in &lt;strong&gt;ComfyUI&lt;/strong&gt;. This ControlNet can influence SDXL such that the generated image “hides” a scan-able QR code, which at first glance, looks like a photo!&lt;/p&gt;</description></item><item><title>Generating animations / videos with SDXL LCM</title><link>https://myByways.com/post/generating-animations-videos-with-sdxl-lcm/</link><pubDate>Mon, 13 Nov 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/generating-animations-videos-with-sdxl-lcm/</guid><description>&lt;p&gt;I never tried generating video clips or animations with SDXL before, simply because my GPU is not powerful enough. But after testing out the &lt;a href="https://myByways.com/post/crazy-fast-image-generation-with-lcm-lora-for-sdxl"&gt;LCM LoRA for SDXL yesterday&lt;/a&gt;, I thought I’d try the &lt;a href="https://huggingface.co/blog/lcm_lora" rel="external"&gt;SDXL LCM LoRA&lt;/a&gt; with &lt;a href="https://github.com/hotshotco/Hotshot-XL" rel="external"&gt;Hotshot-XL&lt;/a&gt;, which is something akin to &lt;a href="https://animatediff.github.io/" rel="external"&gt;AnimateDiff&lt;/a&gt;.&lt;/p&gt;</description></item><item><title>Crazy fast image generation with LCM LoRA for SDXL</title><link>https://myByways.com/post/crazy-fast-image-generation-with-lcm-lora-for-sdxl/</link><pubDate>Sun, 12 Nov 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/crazy-fast-image-generation-with-lcm-lora-for-sdxl/</guid><description>&lt;p&gt;Stable Diffusion keeps improving at an astounding pace! This time, it’s the idea of &lt;em&gt;distilling&lt;/em&gt; a model into a &lt;strong&gt;Latent Consistency Model (LCM)&lt;/strong&gt; for very, very fast image generation with a quality trade-off. On 24 Oct 2023, the distilled &lt;a href="https://huggingface.co/segmind/SSD-1B" rel="external"&gt;Segmind Stable Diffusion 1B (SSD-1B) model&lt;/a&gt; was released, followed by a better implementation in the form of &lt;a href="https://huggingface.co/blog/lcm_lora" rel="external"&gt;Latent Consistency LoRAs for SDXL and SDD-1B&lt;/a&gt; released on 9 Nov 2023.&lt;/p&gt;</description></item><item><title>Using the SDXL-Inpainting 0.1 Model</title><link>https://myByways.com/post/using-the-sdxl-inpainting-01-model/</link><pubDate>Sun, 03 Sep 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/using-the-sdxl-inpainting-01-model/</guid><description>&lt;p&gt;Stability AI just released an new &lt;a href="https://huggingface.co/diffusers/stable-diffusion-xl-1.0-inpainting-0.1" rel="external"&gt;SD-XL Inpainting 0.1 model&lt;/a&gt;. Here is how to use it with &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt;.&lt;/p&gt;</description></item><item><title>Improving faces with Impact-Pack Detailers</title><link>https://myByways.com/post/improving-faces-with-impact-pack-detailers/</link><pubDate>Thu, 31 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/improving-faces-with-impact-pack-detailers/</guid><description>&lt;p&gt;&lt;a href="https://huggingface.co/stabilityai/stable-diffusion-xl-base-1.0" rel="external"&gt;Stable Diffusion XL&lt;/a&gt; has trouble producing accurately proportioned faces when they are too small. Today, I learn to use the &lt;strong&gt;FaceDetailer&lt;/strong&gt; and &lt;strong&gt;Detailer (SEGS)&lt;/strong&gt; nodes in the &lt;a href="https://github.com/ltdrdata/ComfyUI-Impact-Pack" rel="external"&gt;ComfyUI-Impact-Pack&lt;/a&gt; to fix small, ugly faces.&lt;/p&gt;</description></item><item><title>Improving poses with SDXL ControlNets</title><link>https://myByways.com/post/improving-poses-with-sdxl-controlnets/</link><pubDate>Sat, 26 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/improving-poses-with-sdxl-controlnets/</guid><description>&lt;p&gt;I previously &lt;a href="https://huggingface.co/thibaud/controlnet-openpose-sdxl-1.0" rel="external"&gt;tried Thibauld’s SDXL-controlnet: OpenPose (v2) ControlNet&lt;/a&gt; in &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; with poses either downloaded from &lt;a href="https://openposes.com/" rel="external"&gt;OpenPoses.com&lt;/a&gt; or created with &lt;a href="https://github.com/space-nuko/ComfyUI-OpenPose-Editor" rel="external"&gt;OpenPose Editor&lt;/a&gt;. Here are a few more options for anyone looking to create custom poses.&lt;/p&gt;</description></item><item><title>SDXL Revision workflow in ComfyUI</title><link>https://myByways.com/post/sdxl-revision-workflow-in-comfyui/</link><pubDate>Sun, 20 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/sdxl-revision-workflow-in-comfyui/</guid><description>&lt;p&gt;Yesterday, I &lt;a href="https://myByways.com/post/stability-ai-control-loras"&gt;tried out Stability AI’s four Control-LoRAs&lt;/a&gt; but mentioned that I did not understand the output of the &lt;strong&gt;Revision&lt;/strong&gt; “image-mixing” workflow. I’ve since done a bit more experimentation&amp;hellip;&lt;/p&gt;</description></item><item><title>Stability AI Control LoRAs</title><link>https://myByways.com/post/stability-ai-control-loras/</link><pubDate>Sat, 19 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/stability-ai-control-loras/</guid><description>&lt;p&gt;Less than a week after my post testing &lt;a href="https://myByways.com/post/sdxl-1-0-with-sdxl-controlnet-canny"&gt;diffusers/controlnet-canny-sdxl-1.0&lt;/a&gt;, along comes Stability AI’s own &lt;strong&gt;ControlNets&lt;/strong&gt;, which they call &lt;a href="https://huggingface.co/stabilityai/control-lora" rel="external"&gt;Control-LoRAs&lt;/a&gt;! Not one but 4 of them - &lt;strong&gt;Canny&lt;/strong&gt;, &lt;strong&gt;Depth&lt;/strong&gt;, &lt;strong&gt;Recolor&lt;/strong&gt; and &lt;strong&gt;Sketch&lt;/strong&gt; models!&lt;/p&gt;</description></item><item><title>SDXL 1.0 with SDXL-ControlNet: OpenPose (v2)</title><link>https://myByways.com/post/sdxl-1-0-with-sdxl-controlnet-openpose-v2/</link><pubDate>Fri, 18 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/sdxl-1-0-with-sdxl-controlnet-openpose-v2/</guid><description>&lt;p&gt;As &lt;a href="https://myByways.com/post/fooocus-ksampler-custom-node-for-comfyui-sdxl"&gt;promised in my last post&lt;/a&gt;, today I am testing out &lt;strong&gt;Thibaud Zamora&lt;/strong&gt;’s &lt;a href="https://huggingface.co/thibaud/controlnet-openpose-sdxl-1.0" rel="external"&gt;SDXL-controlnet: OpenPose (v2) model&lt;/a&gt; using &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt;. I keep saying I’ll keep my posts short but never do&amp;hellip;&lt;/p&gt;</description></item><item><title>Fooocus KSampler Custom Node for ComfyUI SDXL</title><link>https://myByways.com/post/fooocus-ksampler-custom-node-for-comfyui-sdxl/</link><pubDate>Thu, 17 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/fooocus-ksampler-custom-node-for-comfyui-sdxl/</guid><description>&lt;p&gt;Recently I tried &lt;a href="https://github.com/lllyasviel/Fooocus" rel="external"&gt;Fooocus by Lyumin Zhang (Illyasviel)&lt;/a&gt; which fulfills its promise to allow one to “Focus on prompting and generating” - it is certainly easy to use! But shortly after its release, someone has “ported” the code to &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; as a &lt;strong&gt;Custom Node&lt;/strong&gt;! So of course it’s time to test it out&amp;hellip;&lt;/p&gt;</description></item><item><title>SDXL 1.0 with SDXL-ControlNet: Canny</title><link>https://myByways.com/post/sdxl-1-0-with-sdxl-controlnet-canny/</link><pubDate>Sun, 13 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/sdxl-1-0-with-sdxl-controlnet-canny/</guid><description>&lt;p&gt;The Stability AI documentation now has a pipeline supporting &lt;a href="https://huggingface.co/docs/diffusers/main/en/api/pipelines/controlnet_sdxl" rel="external"&gt;ControlNets with Stable Diffusion XL&lt;/a&gt;! Time to try it out with &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; for Windows.&lt;/p&gt;</description></item><item><title>Scale and Composite Latents with SDXL</title><link>https://myByways.com/post/scale-and-composite-latents-with-sdxl/</link><pubDate>Sat, 05 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/scale-and-composite-latents-with-sdxl/</guid><description>&lt;p&gt;In this post, I experiment with &lt;strong&gt;latent scaling&lt;/strong&gt; and &lt;strong&gt;latent compositing&lt;/strong&gt; with &lt;a href="https://stability.ai/blog/stable-diffusion-sdxl-1-announcement" rel="external"&gt;SDXL 1.0&lt;/a&gt; using &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt;. That is to say, increasing / decreasing the size of the image, and combining multiple images into one à la green screen (chroma key) compositing.&lt;/p&gt;</description></item><item><title>Two Text Prompts (Text Encoders) in SDXL 1.0</title><link>https://myByways.com/post/two-text-prompts-text-encoders-in-sdxl-1-0/</link><pubDate>Thu, 03 Aug 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/two-text-prompts-text-encoders-in-sdxl-1-0/</guid><description>&lt;p&gt;&lt;a href="https://stability.ai/blog/stable-diffusion-sdxl-1-announcement" rel="external"&gt;SDXL 1.0&lt;/a&gt; uses &lt;em&gt;two text prompts&lt;/em&gt; used to guide image generation. In my first post, &lt;a href="https://myByways.com/post/stable-diffusion-sdxl-1-0-with-comfyui"&gt;SDXL 1.0 with ComfyUI&lt;/a&gt;, I referred to the second text prompt as a “style” but I wonder if I am correct. I have no idea! So let’s test out both prompts&amp;hellip;&lt;/p&gt;</description></item><item><title>CLIPSeg with SDXL in ComfyUI</title><link>https://myByways.com/post/clipseg-with-sdxl-in-comfyui/</link><pubDate>Mon, 31 Jul 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/clipseg-with-sdxl-in-comfyui/</guid><description>&lt;p&gt;Onward with &lt;a href="https://stability.ai/blog/stable-diffusion-sdxl-1-announcement" rel="external"&gt;SDXL&lt;/a&gt; and &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt;! Sometimes I want to tweak generated images by replacing selected parts that don’t look good while retaining the rest of the image that does look good. Rather than manually creating a mask, I’d like to leverage &lt;a href="https://github.com/timojl/clipseg" rel="external"&gt;CLIPSeg&lt;/a&gt; to generate a masks from a text prompt.&lt;/p&gt;</description></item><item><title>SDXL with Offset Example LoRA in ComfyUI for Windows</title><link>https://myByways.com/post/sdxl-with-offset-example-lora-in-comfyui-for-windows/</link><pubDate>Sun, 30 Jul 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/sdxl-with-offset-example-lora-in-comfyui-for-windows/</guid><description>&lt;p&gt;Yesterday I mentioned in passing that my &lt;a href="https://myByways.com/post/stable-diffusion-sdxl-1-0-with-comfyui"&gt;Nvidia RTX 2060 with 12GB could not run both SDXL 1.0 Base and Refiner models&lt;/a&gt; in a single &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt; workflow. Today, I show you my workaround and also experiment with adding the &lt;a href="https://stability.ai/blog/stable-diffusion-sdxl-1-announcement" rel="external"&gt;SDXL 1.0&lt;/a&gt; &lt;strong&gt;Official Offset Example LoRA&lt;/strong&gt; to the workflow.&lt;/p&gt;</description></item><item><title>Stable Diffusion SDXL 1.0 with ComfyUI</title><link>https://myByways.com/post/stable-diffusion-sdxl-1-0-with-comfyui/</link><pubDate>Sat, 29 Jul 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/stable-diffusion-sdxl-1-0-with-comfyui/</guid><description>&lt;p&gt;&lt;strong&gt;Stability.ai&lt;/strong&gt; has released &lt;a href="https://stability.ai/blog/stable-diffusion-sdxl-1-announcement" rel="external"&gt;Stable Diffusion XL (SDXL) 1.0&lt;/a&gt; (26 July 2023)! Time to test it out using a no-code GUI called &lt;a href="https://github.com/comfyanonymous/ComfyUI" rel="external"&gt;ComfyUI&lt;/a&gt;!&lt;/p&gt;</description></item><item><title>myByways Simple-SD v1.1 Python script using Safetensors</title><link>https://myByways.com/post/mybyways-simple-sd-v1-1-python-script-using-safetensors/</link><pubDate>Tue, 24 Jan 2023 00:00:00 +0000</pubDate><guid>https://myByways.com/post/mybyways-simple-sd-v1-1-python-script-using-safetensors/</guid><description>&lt;p&gt;&lt;strong&gt;&lt;a href="https://huggingface.co/johnslegers/epic-diffusion" rel="external"&gt;Epic Diffusion&lt;/a&gt;&lt;/strong&gt; recently came to my attention, a high-quality merge of various models by John Slegers: “Epîc Diffusion is a general purpose model based on Stable Diffusion 1.x intended to replace the official SD releases as your default model. It is focused on providing high quality output in a wide range of different styles&amp;hellip;” Figured I’d give it a spin.&lt;/p&gt;</description></item><item><title>Fast Stable Diffusion using Core ML on M1</title><link>https://myByways.com/post/fast-stable-diffusion-using-core-ml-on-m1/</link><pubDate>Sun, 18 Dec 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/fast-stable-diffusion-using-core-ml-on-m1/</guid><description>&lt;p&gt;Recently (around 14 December 2022), Apple’s &lt;strong&gt;Machine Learning Research&lt;/strong&gt; team published &lt;a href="https://machinelearning.apple.com/research/stable-diffusion-coreml-apple-silicon" rel="external"&gt;“Stable Diffusion with Core ML on Apple Silicon”&lt;/a&gt; with Python and Swift source code optimized for Apple Silicon (M1/M2) on Github &lt;a href="https://github.com/apple/ml-stable-diffusion" rel="external"&gt;apple/ml-stable-diffusion&lt;/a&gt;. Here I’m trying it out on a MacBook (though the code also works on iPhones and iPads)&amp;hellip;&lt;/p&gt;</description></item><item><title>myByways Simple-SD v1.0 Python script</title><link>https://myByways.com/post/mybyways-simple-sd-python-script/</link><pubDate>Sun, 09 Oct 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/mybyways-simple-sd-python-script/</guid><description>&lt;p&gt;I refactored my &lt;a href="https://myByways.com/post/adding-clipseg-automatic-masking-to-stable-diffusion"&gt;previous &lt;strong&gt;Stable Diffusion&lt;/strong&gt; code&lt;/a&gt;, to clean up, OO it a little, and add new features like tiling, upscaling, PNG metadata. As I mentioned before, I don’t understand AI/ML&amp;hellip; but I do understand programming! So here is my new, more elegant &lt;strong&gt;Simple-SD v1.0&lt;/strong&gt; Python script.&lt;/p&gt;</description></item><item><title>Adding CLIPSeg automatic masking to Stable Diffusion</title><link>https://myByways.com/post/adding-clipseg-automatic-masking-to-stable-diffusion/</link><pubDate>Wed, 28 Sep 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/adding-clipseg-automatic-masking-to-stable-diffusion/</guid><description>&lt;p&gt;I have more ideas for &lt;strong&gt;Stable Diffusion&lt;/strong&gt;. My nights and weekends are consumed! This time: For inpainting, why create a mask image manually, when A.I. can automatically build a mask from a text prompt? Someone much smarter has already &lt;a href="https://arxiv.org/abs/2112.10003" rel="external"&gt;published a paper (arXiv:2112.10003 [cs.CV])&lt;/a&gt;, with source code, to do just this!&lt;/p&gt;</description></item><item><title>Stable Diffusion script with inpainting mask</title><link>https://myByways.com/post/stable-diffusion-script-with-in-painting-mask/</link><pubDate>Mon, 26 Sep 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/stable-diffusion-script-with-in-painting-mask/</guid><description>&lt;p&gt;More &lt;strong&gt;Stable Diffusion!&lt;/strong&gt; This time &lt;em&gt;attempting&lt;/em&gt; to add inpainting / masking based on my previous code, to merge both &lt;code&gt;txt2img.py&lt;/code&gt; and &lt;code&gt;img2img.py&lt;/code&gt; capabilities, disregarding the out-of-box &lt;code&gt;inpainting.py&lt;/code&gt; code, which does not have parameters for positive or negative prompts. Keyword being &lt;em&gt;attempting&lt;/em&gt;&amp;hellip;&lt;/p&gt;</description></item><item><title>My simplified Stable Diffusion Python script</title><link>https://myByways.com/post/my-simplified-stable-diffusion-python-script/</link><pubDate>Wed, 21 Sep 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/my-simplified-stable-diffusion-python-script/</guid><description>&lt;p&gt;I’ve been playing around with the &lt;a href="https://github.com/CompVis/stable-diffusion" rel="external"&gt;Stable Diffusion&lt;/a&gt; scripts a little (to be exact, &lt;a href="https://github.com/bfirsh/stable-diffusion" rel="external"&gt;Ben Firshman’s version&lt;/a&gt;). To help me understand the script, I decided to re-write it the way I prefer to use it&amp;hellip; either breaking or optimizing it in the process :P&lt;/p&gt;</description></item><item><title>Stable Diffusion image-to-image mode</title><link>https://myByways.com/post/stable-diffusion-image-to-image-mode/</link><pubDate>Sat, 17 Sep 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/stable-diffusion-image-to-image-mode/</guid><description>&lt;p&gt;Following from my previous post, &lt;a href="https://myByways.com/post/ai-generated-images-with-stable-diffusion-on-an-m1-mac"&gt;AI-generated images with Stable Diffusion on an M1 mac&lt;/a&gt;: This time, using the image-to-image script, which takes an input “seed” image, in addition to the text prompt as inputs. In this case the model will use the shapes and colors in the input image as a base for the output AI-generated image.&lt;/p&gt;</description></item><item><title>AI-generated images with Stable Diffusion on an M1 mac</title><link>https://myByways.com/post/ai-generated-images-with-stable-diffusion-on-an-m1-mac/</link><pubDate>Fri, 16 Sep 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/ai-generated-images-with-stable-diffusion-on-an-m1-mac/</guid><description>&lt;p&gt;There has been a lot of buzz about &lt;strong&gt;Stable Diffusion&lt;/strong&gt; for text-to-image synthesis, which saw its Public Release around 22 Aug 22. You can read more on the &lt;a href="https://stability.ai/blog/stable-diffusion-public-release" rel="external"&gt;Stability.AI blog&lt;/a&gt; and try it at &lt;a href="https://huggingface.co/spaces/stabilityai/stable-diffusion" rel="external"&gt;Hugging Face&lt;/a&gt;. What’s groundbreaking is is that is open source, with a pre-trained downloadable model and modest system requirements, so anyone can try it on their own computer&amp;hellip; anyone&amp;hellip; like me!&lt;/p&gt;</description></item><item><title>Running GFPGAN Face Restoration in a container</title><link>https://myByways.com/post/running-gfpgan-face-restoration-in-a-container/</link><pubDate>Mon, 01 Aug 2022 00:00:00 +0000</pubDate><guid>https://myByways.com/post/running-gfpgan-face-restoration-in-a-container/</guid><description>&lt;p&gt;Almost a year ago, Tencent researchers released their &lt;a href="https://github.com/TencentARC/GFPGAN" rel="external"&gt;GFPGAN Face Restoration&lt;/a&gt;, an AI model which is trained specifically on faces, to better upscale and restore details in low-resolution or damaged portrait photos. I thought I’d give it a whirl.&lt;/p&gt;</description></item></channel></rss>