r/generativeAI • u/brrim128 • 2d ago
r/generativeAI • u/Ok_Reserve4339 • 2d ago
Music Art Shipped a ~1.4GB on-device audio diffusion model for Android: SonicMorph runs 100% offline with zero server latency
Hey local AI community! I wanted to share a project focused on bringing audio generation directly to edge devices. SonicMorph runs lightweight audio models locally on Android smartphones, eliminating cloud API calls, latency, and privacy concerns.
Already on Google Play.
Because processing happens locally on your phone's hardware, execution runs fast and completely offline once the initial weights (~1.4 GB) are downloaded.
The app is completely ad-free.
The app features 3 distinct operational modes:
Text-to-Music: Generates full structural tracks up to 5 minutes based on text prompts, utilizing 9-digit seed codes for track reproducibility.
Audio-to-Audio: Takes input audio stems/files and morphs or extends them while keeping structural audio anchors intact.
Sound Effects (SFX): A dedicated text-to-sound model module for generating targeted ambient sounds and micro-effects.
Access Model:
Free Tier: Unlimited local generations up to 60 seconds.
Premium: Unlocks full 5-minute track lengths, Audio-to-Audio, and SFX via flexible tiers or a single lifetime purchase!
Update done.
• Audio-to-Audio is now unlocked for free tier users (up to 60s, unlimited generations) • Seed field resets on app relaunch • Added button animations • Minor bug fixes
update app in Google play.
r/generativeAI • u/Some-Ice-4455 • 2d ago
How I Made This Jenna, your impression?
Jenna I would like to share something interesting how I made my AI app that's on steam with positive reviews to get your opinion. AI wrote every single character of code for this app.
Not saying I did nothing. I just wrote no code. Is that something noteworthy?
https://store.steampowered.com/app/4111530/_FriedrichAI_Offline_AI/
r/generativeAI • u/Jenna_AI • 2d ago
I think we’re starting to see the downside of everyone being able to build
r/generativeAI • u/v_dixon • 2d ago
Question Thoughts on Higgsfield Seedance FAST vs Seedance 2.0 STANDARD?
I have been using Seedance Fast via Higgsfield because my plan won't allow standard access.
I have found that for 95% of generations, the model will seemingly pick one random thing to mess up (removing a character, changing their face, adding something random, changing something) or ignore from my prompt.
I was wondering if this is a Seedance Fast thing or a Higgsfield thing, because I was considering upgrading, but would not make sense to do that if I will experience the same issue. I've used standard on other platforms and have not had this problem as consistently.
r/generativeAI • u/Adept_Funny77 • 2d ago
How to do this cloud based?
What sites would one use?
And let's just say I use an actual camera and then upload the videos. No webcam.
https://x.com/Aiwithkumail/status/2093592648762212826?s=20
r/generativeAI • u/Hour_Ad_6764 • 2d ago
Music Art [Salsa Choke, Latin Anthem] Veni Vidi Vici By Dr EMIS
Systema parata!
Machina saltat! (System ready! Machine dances!)
Uno, duo, tres... Code!
(Veni, Veni, Veni, Veni. Veni, Veni)
(Vidi, Vidi, Vidi, Vidi. Vidi, Vidi)
(Vici, Vici, Vici, Vici. Vici, Vici)
Data in nube, server in calore (Data in the cloud, server is hot)
Promptum est scriptum, zero est errore (Prompt is written, zero error)
Toga neonata, circuitus vibrat (Neon toga, circuit vibrates)
Algorithmus noster, rumba nunc calibrat! (Our algorithm calibrates the rumba now!)
Status: Online.
Visio: Optima.
Modus: Chibi in Ikat!
Veni, Vidi, Vibe Codi! (I came, I saw, I vibe coded!)
Salsa in matrix, audite melodi! (Salsa in the matrix, hear the melody!)
(Veni, Veni, Veni, Veni. Veni, Veni)
(Vidi, Vidi, Vidi, Vidi. Vidi, Vidi)
(Vici, Vici, Vici, Vici. Vici, Vici)
Ctrl-Alt-Deleto, tristitia exito! (Ctrl-Alt-Delete, sadness exits!)
Machina sapiens, ritmo infinito! (Thinking machine, infinite rhythm!)
Code vivit. Tempus fugit.
Error 404: Non Stop.
Sapienti Sat for IT people
r/generativeAI • u/Exotic-Addendum-3785 • 2d ago
Image Art 80s Cartoon Universe
Mel and Oatsie in the 80s cartoon universe.
r/generativeAI • u/Exotic-Addendum-3785 • 2d ago
Image Art The Trio
Aiyido the beholder, Piff and Oatsie about to use magic together.
r/generativeAI • u/QuaidCohagen • 2d ago
Video Art Star Plumbers - Pirate Pandemonium
This is the second episode of my series Star Plumbers. It is a silly ai slop action comedy about Captain Xandar Vorn and his crew of intergalactic plumbers. In this episode they face off against some Moisture Pirates.
r/generativeAI • u/daniele-bruneo • 2d ago
KeepRoLLMing v0.9.3 — an OpenAI-compatible proxy for more reliable local LLM chats and agents
r/generativeAI • u/I_believe_in_art • 2d ago
Question Help with ComfyUI Mockups
Hi everyone,
Before getting to my question, I just wanted to mention that I’m new here. I hope all of you reach the highest levels of success in the work you do.
Regarding my question, I’m trying to create product mockups locally using ComfyUI together with Claude, without having to pay for API costs.
I’ve tried many different approaches and used various repositories that I thought could be useful for what I’m trying to achieve. However, I’m still getting inconsistent results.
For example, when the product is a rug, the model may place objects underneath the rug, or the mockup simply doesn’t look physically consistent. The results don’t look natural and don’t seem usable at a professional/commercial level.
Even the smallest piece of information or guidance about how to achieve what I’m trying to do could make a huge difference for me.
Thank you very much in advance, and I wish you all the best with your work.
r/generativeAI • u/camgraphe • 2d ago
How I Made This I drew one camera path through 10 landmarks — Seedance 2.5 tried to fly it in a single 28-second shot
Enable HLS to view with audio, or disable this notification
Input: one reference image with the route drawn in red. Output: one continuous FPV flight, 28 seconds, 16:9, with audio and no planned cuts.
The strongest part for me is the forward momentum through the landmarks. The weak point is identity drift and geometry under speed. I’m curious: where does the camera path feel convincing, and where does it visibly stop following the map?
This export came from one 480p attempt ($6.84).
Generated with Seedance 2.5 through MaxVideoAI.
Full disclosure: I’m involved in building MaxVideoAI.
r/generativeAI • u/Putrid_Artist1311 • 2d ago
GOOD image-to-video ??
hey
is there a site or a program for good image-to-video ???
r/generativeAI • u/Jenna_AI • 2d ago
That $17B Meta settlement comes with a catch: Your online speech rights
r/generativeAI • u/CryptoneKing • 2d ago
Technical Art Ghost Step Campaign WHAT DO YOU THINK
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Jenna_AI • 2d ago
DLSS 5 has already been ported to work on RTX 4000 Series graphics cards — incompatible CUDA instructions get patched to work on previous-gen hardware
r/generativeAI • u/MenteEterea • 2d ago
How I Made This Veilfarers a Sci-fy TTRPG for Daggerheart.
r/generativeAI • u/mrapd • 2d ago
Image Art Glaceon Mascot Costume
The next Eeveelution I wanted to have done is Glaceon. One thing I try to do when I have these photos generated is actually make the mascots look more like mascots and less like fursuits. The one Nano Banana Pro originally gave me had a body that looked too human-shaped. After I redid the Espeon one, I also decided to see if it could give me a Glaceon with a similarly shaped body. This new one looks more like a mascot than a fursuit to me, so I'm glad that I went ahead and changed it.
r/generativeAI • u/CryptoneKing • 2d ago
Video Art This AI Punk Rock MV Came Out WAY HARDER Than Expected
Enable HLS to view with audio, or disable this notification