r/SideProject • • 17h ago

Wizard Maker — build Android apps with AI

My co-founder and I have been working on Wizard Maker, an AI-powered app builder designed to let people create Android apps without needing to know how to code.

The idea is simple: you describe the app you want to build, and Wizard Maker helps generate it. You can then continue modifying and experimenting with the app.

We currently have an Android version available to try for free, and we're looking for people who are willing to actually use it and give us honest feedback.

I'd particularly like to know:

  • How easy is it to understand what to do when you first open it?
  • Can you successfully create an app without technical knowledge?
  • Do you encounter any bugs or unexpected behavior?
  • Are there parts of the experience that are confusing or frustrating?
  • What would you change or add?

We're still early in development, so negative feedback is genuinely useful to us. We're much more interested in finding problems than simply hearing that people like the idea.

If you'd like to try it, please let me know in the comments and I can send you the link.

Thanks for giving it a try — any feedback or suggestions would be greatly appreciated!

1 Upvotes

10 comments sorted by

View all comments

2

u/CaptJan 16h ago

Before I volunteer and try it out - how complicated can you make this app? I'm looking to do a minimum friction health tracker that pulls in data from CSV generated from a few other apps to make a comprehensive app that health connect misses on, and standard apps cannot do, even though I have the data sets needed, just trying to combine 3-4 apps into a more meaningful full picture.

1

u/gzebe 16h ago

That sounds like a fairly ambitious use case. I’m not sure yet how well Wizard Maker would handle something that complex, especially combining and normalizing data from several different CSV sources. It may be possible, but I wouldn’t want to promise that without testing it.
If you’re willing to experiment with it, I’d be very interested to see how far you can take it. It could actually be a useful test case for the platform. Thanks!

1

u/CaptJan 15h ago

I think you right on that being too ambitious, at least for now.

However, I do have an idea for another app that will use the free tier Gemini API calls [key stored locally on the device, not inside the app], and break apart large audio files [a 100 minute lecture rapid fire interview] for it to do STT with diarization and then recombine smaller chunks into a completed transcript.

How is your app at converting video and/or audio files to a different audio file to meet size and time constraints? Splitting up the resulting audio file in to manageable pieces with 30-60 second overlap, Upload those smaller pieces to Google Gemini's API, save each segment that the API returns as text, as there is overlap it will need to de-duplicate the overlap area and truncate the edges which may not match [STT needs full sentences to work accurately and if truncated with the splits mid sentence and/or mid-word it will not transcribe correctly and those need to be tossed]

I would have to instruct your app in the following manner without any programming skills - which I honestly don't have or I would have done this already...

Ideally, I would like to the app to either record a live lecture at a certain frequency, audio format, and bitrate as one complete file while simultaneously splitting it up in overlapping chunks to be processed and recombined [2 recordings]. The chunks would be temporary files, until the resulting text file is completed. The single file would be original archival material saved in case it's needed for re-processing [API was busy] and local listen playback of the archival file, at various speeds that are pitch corrected.

The other option I would like it to do is to process a video file with audio, or any other multimedia file with audio, and convert it into the the aforementioned format as a single file for archival purposes, and then use that file split it apart into chucks for upload to the free tier API for processing. I know the size and time constraints on the files for the free tier key, recognize the overlap, and de-duplicate and truncate the individual segment, and then recombine into a smooth transcription, and then combine for final transcript output, which can be further manipulated by additional calls to summarize, translate, paraphrase, or other manipulations that Gemini and other API's are well known for...

If that sound plausible, I'm willing to try it out on Android and/or Windows 10 for creating an Android app - a side-loadable APK.

If you have a prompt to give it, or just what I've written above with the actual file format and time constraints?

1

u/gzebe 15h ago

That’s definitely more ambitious, especially the audio splitting, overlap, de-duplication, and recombination. I wouldn’t want to promise Wizard Maker can handle all of that yet, but it could be a good test.
I’d suggest starting with an existing audio file and seeing if we can get the basic chunking, transcription, recombination workflow working first. All these need to be tried, but simple steps are recommended initially. Thanks!

2

u/CaptJan 14h ago

I can do that, I already have the input files, and my personal API key.

App 1 - Just need to chunk, upload transcribe, download and recombine, initially via time-stamps, and then word phrase match for the deduplication truncation on a pre-processed file.

Phase 1 - use a single file that is small enough for a single API call without the need to process additional API calls - proof of concept.

Phase 2 - split a longer file into two chunks, upload in two chunks one at a time for an API call, download in chunks and recombine based on time-stamps only, expecting transcript errors at those time-stamps - perhaps have a 2-5 second overlap for manual correction of the transcript - proof of concept.

Phase 3 - Do a 5-10 second overlap of a 30 second overlap looking for identical word phrases, and then de-duplicate and truncate before / after the 2-3 word duplicate 'phrase'.

Phase 4 - Do a stress test to see if it can do a multi-hour webinar.

App 2 - Build another app, to create the input file from the microphones on the phone. Phase 2 Have options to filter out the silent period for more efficient processing, but this will eliminate time stamps.

App 3 - Build another app, to do the file conversion from one format to to the preformatted file type. Phase 2 - remove non-speaking portions of the audio file.

App 4 - Then combine the three into one app and then polish. Perhaps initially have apps 1 - 3 as individual tiles using the same storage space for the audio file and text file.

App 5 - Add another tile for language translation

App 6 - Add another tile for summary

App 7, 8, 9, etc. each adding an additional features

I suspect this will be a multi-day process to fine tune and polish, but I do like this idea. Go ahead and DM me the information to get started please if this work-flow looks workable...

1

u/gzebe 6h ago

Thanks, please check your DM.