Recent Posts

Pages: [1] 2 3 ... 10
1
Almost forgot.  Minimax H3 is surprisingly good at the entranced voice.  The problem is pacing.  If you let Minimax handle everything, you'll get those pregnant pauses AI is famous for.
I want to edit the audio separately, so the pacing is good.  Then offer the audio as reference. 

If Minimax really is best for this, I'll simply render scenes with minimal video, then use that audio.

It turns out Minimax H3 really was best at producing a hypnotized voice I like.  But, quality was bad.  Worse, it was inconsistent.  I got maybe 3 lines out of it, just enough to get my hopes up, and then it started producing a different voice.  Sometimes male.  Yuck! ;)  Here's how I fixed it.
1) Quality improved incredibly by adding an Audio Preview node and routing the audio there instead of using the MP4 output.  You can save to flac from the Audio Preview node.  Also, I used an image size of 96x96 to keep render times down.   32x32 produced low quality audio and weird results.
2) Once H3 gives you something you like, you can get consistency by using the clone node variant of one of the TTS's.  I got the best results with Chatterbox.  It adds a little bit of emotion, but not much.  Plug the same training audio into DramaBox when you want normal, or extreme, emotions.  Interestingly, Chatterbox has an "Exaggeration" value that can produce some pretty wild results.  Keep it low for your hypnotized friends.

Lip sync is next for me.  The key seems to be notifying Minimax H3 which audio reference input is being used and provide it with a transcript.  Otherwise you get gibberish that sounds vaguely like the audio you provided.

I can see this process frustrating me.  It is a LOT of work.  Dialog is generated about a line at a time.  Then I need to edit it in Audacity so the pacing is right.  Then I need reference images for each key moment in each scene.  Then I need to arrange the script with stage directions for H3.  And lots of re-rendering when it doesn't work.  So far, I'm having fun, though.

Any workflows you would recommend?

Sorry it took me so long to respond.   I haven't made much progress in the last month.  Pixaroma is my favorite source on youTube for workflows.  His work well and he does a good job of  explaining how and why things work.

At this point getting an entranced look is fairly easy for me.  Sometimes she'll want to respond to what's going on around her, but you can usually prompt "her" out of it.

The problem I have now is the induction.  I want a momentary subtle realization that something isn't right, followed by a slow loss of all tension and fear.  H3 makes her give up and look bored.   Sometimes, if you don't give your characters things to say, they'll speak gibberish.

Anyway, this is fun.  I can leave my PC alone for 10 minutes churning out 10 seconds of video.  I'll change the prompt a bit and come back in 10 minutes (30-40 minutes if in HD).
2
Is there any chance you want to share a link to your content to showcase?
I'm hoping to start a thread here soon! I have many hundreds of pictures, but I'd have to comb through them to find some good 'standalone' images that suggest a self contained story

I look forward to it. Hopefully I check in to get notified of it, but if not if you message me through the forum I would love to check it out.
3
Is there any chance you want to share a link to your content to showcase?
I'm hoping to start a thread here soon! I have many hundreds of pictures, but I'd have to comb through them to find some good 'standalone' images that suggest a self contained story
4
I would point out that GPT is a lot less prudish than it used to be, and as long as you tell it all the sex scenes are 'adults making out', its willing to let you take it in some pretty spicy directions (provided you are willing to edit and control the flow paragraphs at a time, and write the spicy scenes yourself or with another model).
I have noticed in general AI seems to be getting a bit better at allowing more sexual/suggestive material - hopefully this is a sign that a happy medium is slowly being reached between "completely bland & sexless, never spicier than a workplace motivational poster from the 1990s" and "you can go ahead & do anything, even the most depraved illegal material is fine, surely there won't be any blowback from this".

Update: it seems that within the last 24h Google Flow has updated to Nano Banana 2.1 (and it might be my imagination, but I feel like the usage limit is much higher - although this might just be temporary, or perhaps there was a mid-day reset when they did the update). I haven't played around with it too much but it seems subtly better at getting the proportions between characters & objects right, and at replicating various film styles. The big improvement I notice is that it seems way better at posing characters naturally in the environment in a wider variety of ways. In the past there was a tendency to default to the old "character standing in a neutral pose in front of a background that they don't interact with, which may as well just be a painted wall". Now they're much more likely to be interacting with nearby objects in organic ways, and the lighting on characters feels much more cohesive with the surrounding environment.

The other change I've noticed (which, again, might just be my imagination) is that the new model seems less likely to reject prompts, choosing instead to simply ignore the most offensive elements. So, the pro here is that you'll get more images generated (rather than just wasting tokens), and the con is that perhaps you'll have less success occassionally tricking the AI into giving you something truly spicey. In general, in terms of what it'll allow, it seems comfortably PG-13, although if one were to start leaning very heavily into certain kinds of fetishistic language it might push back on it a bit.

Is there any chance you want to share a link to your content to showcase?
5
I would point out that GPT is a lot less prudish than it used to be, and as long as you tell it all the sex scenes are 'adults making out', its willing to let you take it in some pretty spicy directions (provided you are willing to edit and control the flow paragraphs at a time, and write the spicy scenes yourself or with another model).
I have noticed in general AI seems to be getting a bit better at allowing more sexual/suggestive material - hopefully this is a sign that a happy medium is slowly being reached between "completely bland & sexless, never spicier than a workplace motivational poster from the 1990s" and "you can go ahead & do anything, even the most depraved illegal material is fine, surely there won't be any blowback from this".

Update: it seems that within the last 24h Google Flow has updated to Nano Banana 2.1 (and it might be my imagination, but I feel like the usage limit is much higher - although this might just be temporary, or perhaps there was a mid-day reset when they did the update). I haven't played around with it too much but it seems subtly better at getting the proportions between characters & objects right, and at replicating various film styles. The big improvement I notice is that it seems way better at posing characters naturally in the environment in a wider variety of ways. In the past there was a tendency to default to the old "character standing in a neutral pose in front of a background that they don't interact with, which may as well just be a painted wall". Now they're much more likely to be interacting with nearby objects in organic ways, and the lighting on characters feels much more cohesive with the surrounding environment.

The other change I've noticed (which, again, might just be my imagination) is that the new model seems less likely to reject prompts, choosing instead to simply ignore the most offensive elements. So, the pro here is that you'll get more images generated (rather than just wasting tokens), and the con is that perhaps you'll have less success occassionally tricking the AI into giving you something truly spicey. In general, in terms of what it'll allow, it seems comfortably PG-13, although if one were to start leaning very heavily into certain kinds of fetishistic language it might push back on it a bit.
6
AI Content / Re: Apropos AI Art
« Last post by apropos on October 06, 2026, 10:02:42 am »
Damn, I'm sorry to hear that. And thank you for the support while you were a member, I really appreciate it.

Have you tried appealing via https://www.deviantart.com/contact-us?

If you think it's mature content that you accidentally left outside of the paywall and/or was untagged mature content and/or was a celebrity and/or didn't have a "this is consensual fiction" warning it might be worth asking if they could let you know what the problem was so you can rebuild from scratch and make sure it doesn't happen again.

Genuinely don't know if that would work, but might be worth a try.
7
I would point out that GPT is a lot less prudish than it used to be, and as long as you tell it all the sex scenes are 'adults making out', its willing to let you take it in some pretty spicy directions (provided you are willing to edit and control the flow paragraphs at a time, and write the spicy scenes yourself or with another model).
8
AI Content / Re: iskander-zombie has created some Downing street related images
« Last post by Svengli on October 06, 2026, 01:12:19 am »

It just wanted to say that I really enjoyed the Child Psychology image series.
9
For AI fiction I usually like to generate about 5-10 paragraphs at a time, then refine and choose the next direction. When making something like that, would my puny 8gb 4060 manage? I've never used local LLMs.
5-10 paragraphs is pretty normal for roleplay chats, but gets a little tedious if you want to write something longer or more complex, unless you nail the generation each time with your prompts. (most AIs will lose coherency faster with larger numbers of prompts, and not with longer generation) I tend to regenerate a few times, so it's easier to just do a longer block and edit manually afterwards.

Well, I'm running on a 5700XT and letting it overflow onto the 5800X CPU with 32GB RAM, and it's enough to run Q4_K_M with 64k context at Q8_0 cache quantization. It's slow, (~15t/s prompt, ~2t/s generation) but it work, and it's completely private. I'm upgrading the GPU though, 9060XT 16GB, so that should make it possible to put the entire model into VRAM across both cards, which should improve speed massively. (I'll keep the context in RAM though, so I can run a bigger model with better outputs)
10
Looks like all the suggestions so far have been for images, video, and audio and not really anything for text, so I'll give my recommendations for a local LLM that can help with text gen, and roleplay.

First, get https://github.com/thomas9120/LLama-GUI and install it. This is what allows you to run and interface with the LLM.
Then get the LLM. I recommend https://huggingface.co/DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF if you have less than a 5090 to run it. (lower than IQ4_XS quantization is not recommended)
Next, get https://github.com/SillyTavern/SillyTavern for the primary interface. It is basically roleplay chat, and you can set up various different characters, world rules, environments, and all sorts of things. It also seems to help bypass the few remaining censorship refusals that the LLM model I recommended has through basic jailbreaking techniques. (you can download several characters and lorebooks here https://character-tavern.com/ )

I highly recommend generating no more than 40 paragraphs at a time, and refining your prompt each time to get the output you actually want, or close enough to allow manual editing. (though if your prompt is long and detailed enough, you can produce as much as 100 paragraphs in a single go)

For AI fiction I usually like to generate about 5-10 paragraphs at a time, then refine and choose the next direction. When making something like that, would my puny 8gb 4060 manage? I've never used local LLMs.
Pages: [1] 2 3 ... 10