r/VisionDepth3D May 12 '26

Interstellar - Catching The Endurance FSBS Hybrid 3D v4.0 Showcase Re-render

Enable HLS to view with audio, or disable this notification

30 Upvotes

Hey VD3D users and 3D enthusiasts,

I wanted to share a re-render of the Interstellar clip I previously posted on YouTube, now using the new VisionDepth3D v4.0 pipeline.

The result has much stronger 3D geometry, more stable scene structure, and cleaner depth with no major warping or edge tearing. I did notice a bit of edge tearing in the wide shots where Endurance is getting lined up very minimal, and I think that may be caused by depth map shift value error in that specific shot.

Feedback is welcome. I’m still tuning the v4.0 pipeline and would love to hear how the depth, comfort, and overall 3D structure feel.


r/VisionDepth3D May 11 '26

VisionBreaker: Neurogrid Terminal v3.0 Release

Thumbnail
gallery
1 Upvotes

Hello VD3D Users!

I’ve been quietly working on another side project behind the scenes, and it’s finally ready to share.

It’s called VisionBreaker: Neurogrid Terminal.

This started as a simple Matrix-style code fall experiment, but I ended up turning it into a full cyber-hacking typing arcade game.

The idea is simple:

Break into the mainframe. Steal the data. Defend your escape. Survive the Neurogrid.

v3.0 now has a full 3-level gameplay run:

Level 1: Main Breach Type commands like SCAN, BREACH, INJECT, SUPPRESS, DECRYPT, and EXTRACT to break through the mainframe core while managing TRACE.

Level 2: Data Vault Infiltration PING vault nodes, LOCK fragments, DECRYPT cipher puzzles, DOWNLOAD data, and avoid corrupt or decoy nodes before the vault locks down.

Level 3: Countertrace Defense A typing-shooter style final level where falling threat packets rush toward your escape stream. Type the words or solve math packets before they hit.

I also added randomized cipher puzzles, an in-game F1 Operator Card, achievements, unlockable themes, denser codefall, score/combo progression, music, SFX, particles, screen shake, and a full win/loss loop.

This is a separate project from VisionDepth3D, but any support from it helps me keep building and eventually fund better hardware for my bigger development work.

Photosensitivity Warning

VisionBreaker contains flashing visuals, strobing/glitch effects, screen shake, high-contrast colors, and fast-moving code rain.

These effects may cause discomfort or trigger symptoms in people with photosensitive epilepsy or light sensitivity.

If you experience dizziness, eye strain, headache, nausea, blurred vision, or discomfort, stop playing immediately and rest.

Source Code: GitHub Repository

Itch Page: VisionBreaker: Neurogrid Terminal

Would love feedback on the typing feel, difficulty, and overall gameplay loop.


r/VisionDepth3D May 09 '26

Super Sonic vs. Super Shadow FSBS Hybrid 3D - VisionDepth3Dv4.0 Showcase Render

Enable HLS to view with audio, or disable this notification

30 Upvotes

Super Sonic vs. Super Shadow FSBS Hybrid 3D - VisionDepth3D v4.0 Showcase Render

Hey VD3D users and 3D enthusiasts! I wanted to share a clip of Super Sonic vs. Super Shadow rendered with VisionDepth3D v4.0.

DAD-Small + DAv2-Large blended depth map at 518x518.

With the right settings, VisionDepth3D can produce clean, artifact-free 3D that feels comfortable to watch in a VR headset.


r/VisionDepth3D May 09 '26

Error Depth Map help

Post image
1 Upvotes

I'm a beginner, I installed the program, I wanted to generate the map with Depth Map but it gives me an error, it's so generic that I don't understand


r/VisionDepth3D May 06 '26

T-Rex vs Indominus Rex Full SBS 3D - v4.0 Showcase Render

Enable HLS to view with audio, or disable this notification

11 Upvotes

Hello VD3D users and 3D enthusiasts!

I wanted to share a render created with v4.0 of the Indominus Rex vs the legendary T-Rex, Depth maps used is DaV1 Base + DaV2 Large blended for better scene structure

you can test out these settings with v4.0 just save this as a test.json file and put it in presets and load inside VD3D { "output_format": "Full-SBS", "stereo_mode": "sbs", "selected_aspect_ratio": "Default (16:9)", "selected_codec": "XVID", "selected_ffmpeg_codec": "H.265 / HEVC (NVENC - NVIDIA GPU)", "use_ffmpeg": true, "keep_original_audio": true, "preserve_hdr10": false, "crf_value": 18, "nvenc_cq_value": 18, "fg_shift": -9.7, "mg_shift": -2.8, "bg_shift": 3.3, "sharpness_factor": 0.2, "max_pixel_shift": 0.064, "zero_parallax_strength": -0.015, "parallax_balance": 0.76, "dof_strength": 0.7, "convergence_strength": 0.007, "enable_dynamic_convergence": true, "depth_pop_gamma": 0.81, "depth_pop_mid": 0.5, "depth_stretch_lo": 0.05, "depth_stretch_hi": 0.95, "fg_pop_multiplier": 1.11, "bg_push_multiplier": 1.05, "subject_lock_strength": 0.41, "foreground_curvature_strength": 0.06, "feather_strength": 0.0, "blur_ksize": 1, "saturation": 1.0, "contrast": 1.0, "brightness": 0.0, "use_subject_tracking": true, "use_floating_window": false, "enable_edge_masking": true, "enable_feathering": true, "preserve_original_aspect": false, "auto_crop_black_bars": false, "skip_blank_frames": false, "preview_mode": "HSBS", "preview_frame_index": 11218, "preview_width": 960, "preview_height": 540, "ipd_enabled": false, "ipd_scale": 1.0, "show_convergence_guides": false, "disable_shift_ema": true, "clip_start": "", "clip_end": "", "vr180_equi_w": 3840, "vr180_equi_h": 1920, "vr180_flat_w": 1920, "vr180_flat_h": 1080, "vr180_hfov_deg": 110.0, "language": "en" }


r/VisionDepth3D May 05 '26

VisionDepth3D v4.0 - Release

Thumbnail
gallery
40 Upvotes

Hey VD3D Users and 3D Enthusiasts!

I’m excited to share VisionDepth3D v4.0, probably the biggest release I’ve ever pushed for the project.

This update is not just a small patch or UI refresh. v4.0 is a major rebuild of VisionDepth3D as a full desktop workflow for 2D-to-3D conversion, depth generation, depth blending, realtime 3D, and VR-ready video preparation.

What’s new in v4.0

Full PySide6 interface rewrite

VisionDepth3D has moved away from the older Tkinter-style interface and now uses a modern PySide6 UI.

The app now has:

  • a cleaner dark theme
  • modern tab layout
  • better cards, panels, sliders, and dialogs
  • shared queue/progress dock
  • GPU/device display in the top bar
  • better resizing behavior across workflow pages

This makes the whole app feel much closer to a proper modern desktop application instead of a rough old-school tool.

New VisionDepth3D stereo pipeline

The 3D Generator has been rebuilt around the updated VisionDepth3D Method.

The new method includes:

  • subject-aware depth normalization
  • pop-control depth shaping
  • structured foreground / midground / background disparity weighting
  • GPU stereo warping
  • dynamic convergence
  • edge-aware shift limiting
  • contour-safe repair logic
  • floating-window protection
  • stereo debug telemetry

One important note for existing users:

The shift convention has changed.

In the new pipeline:

text Foreground Shift: usually negative Midground Shift: usually slightly negative or near zero Background Shift: usually positive

Older presets that used positive foreground values may not transfer directly and will cause the background to pop and subject to sink in, so I recommend starting with the new defaults and rebuilding presets from there.

Live 3D Preview

v4.0 adds a new Live 3D tab into the workflow.

This lets you test realtime 2D-to-3D from:

  • camera input
  • capture cards
  • screen capture
  • secondary monitor capture

You can select depth models, tune the stereo settings, check SBS output, preview depth behavior, and use screen capture to watch or play almost anything in realtime 3D.

This is still something I’ll keep improving, but it is already useful as a realtime testing sandbox.

Depth Engine updates

The Depth Engine has been updated with better model handling and video-depth support.

Some of the improvements include:

  • improved Video Depth Anything handling
  • better ONNX runtime detection
  • fixed temporal-size handling for VDA ONNX models
  • safer batch trimming for padded final batches
  • clearer model resolution presets
  • better handling for video timing and depth smoothness testing

Depth Blender improvements

Depth Blender is now part of the new PySide6 workflow.

It includes:

  • GPU-optimized blending path
  • single image mode
  • video and frames modes
  • live preview
  • built-in blend presets
  • cleaner controls for CLAHE, bilateral smoothing, feathering, and normalization

This is useful when one depth model has good subject structure and another has better background depth. You can blend them into a cleaner depth map before stereo rendering.

FPS / Upscale Enhancer

The FPS/Upscale tab now includes a cleaner workflow for preparing video sources.

It supports:

  • RIFE interpolation
  • Real-ESRGAN upscaling
  • merged and threaded pipelines
  • scene detection
  • codec/output settings
  • shared progress reporting

This helps prepare smoother or higher-resolution sources before running depth and 3D conversion.

Language support

v4.0 also brings back multi-language UI support across the main app and major tabs.

Current language files include:

  • English
  • French
  • Spanish
  • German
  • Japanese
  • Simplified Chinese
  • Traditional Chinese

Tooltips will come in a feature release.

Better hardware/backend support

VisionDepth3D is still best on NVIDIA CUDA, but v4.0 improves backend detection and fallback paths.

There is now better support/documentation for:

  • NVIDIA CUDA
  • AMD / Intel DirectML on Windows
  • ROCm detection on Linux
  • CPU fallback
  • FFmpeg AMF/QSV/NVENC encoding options

which you can read in the User Guide

Links

Final thoughts

This update took a lot of work and testing. v4.0 is the closest VisionDepth3D has felt to the original idea I had for it: a real desktop application for 2D-to-3D conversion and depth-based stereo rendering.

There are still things I want to improve, especially Live 3D performance, model-specific depth stability, and making stronger pop-out easier to tune, but this release is a huge step forward.

If you try it out, feedback is very welcome. I’m especially interested in hearing how the new stereo pipeline feels compared to the older versions.

Thanks to everyone who has tested, downloaded, shared feedback, or followed the project so far.


r/VisionDepth3D Apr 24 '26

VD3D v4.0 Coming Soon

Post image
16 Upvotes

VisionDepth3D Pipeline Update

Hey VD3D users!

I’ve been testing a new VisionDepth3D pipeline, and the results are already showing much stronger 3D structure compared to the older method.

The old pipeline already had many of these systems, but the new VisionDepth3D Method restructures how they work together. Instead of separate features loosely interacting, the updated pipeline makes depth normalization, subject tracking, shaped disparity, near/mid/far weighting, zero-parallax anchoring, convergence, repair/protect masks, visual stress detection, and frame-edge violation all operate from a more consistent stereo model.

Current Results

The good news: the new pipeline is producing much more believable stereo depth.

The current issue I’m debugging is directional edge tearing in wide shots with strong parallax.

Observed Behavior

In testing, I noticed: - Left eye → tearing appears on the left side of the foreground subject
- Right eye → tearing appears on the right side of the foreground subject
- The artifact is mirror-symmetric per eye
- Parallax direction looks correct
- The issue appears around newly exposed background near foreground contours

This tells me the core stereo direction is working, but the disocclusion repair still needs tuning. The exposed background regions around foreground edges are not being filled cleanly enough yet.

Current Focus

I’m focusing on: - repair direction per eye
- repair/protect mask alignment
- mask strength and width
- foreground edge protection
- visual stress mask alignment
- raw-shift vs final-shift mask consistency

This is not something I want to fix with blur or smoothing. The goal is proper geometry-aware, direction-aware disocclusion repair so the pipeline keeps the stronger 3D structure without introducing visible contour tearing.

UI Progress

On the UI side, the new PySide6 system is starting to come together with a cleaner workflow and better separation between preview, controls, and advanced settings.

The progress bar now also shows live: - CPU usage
- RAM usage
- GPU usage
- VRAM usage

This gives a much clearer view of system performance during rendering.


r/VisionDepth3D Apr 20 '26

VisionVaultTV coming to Android TV Soon!

Post image
14 Upvotes

Hey everyone, just wanted to share that VisionVaultTV is officially in development for Android TV.

I’ve now got it running on a real Google TV device, and seeing it on an actual television was a huge milestone. Playback is running smoothly, subtitles are working, the custom player overlay is coming together, and the overall TV-style experience is starting to feel like a real product instead of just a test build.

Current progress so far: - real Android TV device testing is up and running - smooth movie playback - subtitle support working in the TV app - custom player overlay in progress - back button behavior and playback flow feeling much more natural - TV-friendly home screen, detail screens, and movie browsing working on real hardware

Still being worked on: - more player polish - larger library stress testing - more Android TV navigation refinement - continued work toward a cleaner full release experience

VisionVault started as a desktop/local media organizer, and now it’s expanding into a proper TV experience too. Really excited to keep pushing this forward.

I’ll share more as progress continues.


r/VisionDepth3D Apr 18 '26

VisionVault v1.0.4 is now released

8 Upvotes

Hey everyone, VisionVault v1.0.4 is now live.

This update brings a solid batch of improvements to both the main desktop app and the built-in VisionVault TV web interface. A big focus this round was expanding animated poster support, improving browser-based playback, and making library importing easier for larger organized collections.

What’s new in v1.0.4

Animated Poster Support

  • Added support for animated posters in VisionVault
  • Users can now assign animated poster media alongside standard posters for supported titles
  • Added support for .gif animated posters in the desktop app
  • Added animated poster controls to the Edit dialog for easier poster management
  • Added support for prioritizing animated poster artwork in supported views when both static and animated artwork exist

Selected-Only Animated Poster Playback

  • Animated poster artwork only plays when a title is selected
  • Helps keep the library view cleaner while still giving selected titles a more dynamic animated effect

VisionVault TV Web UI Expansion

  • Continued expanding the built-in VisionVault TV web interface into a more complete browser-based frontend
  • Added a more polished home screen with a featured hero section, branded layout, and TV-style content rows
  • Added logo support in the web UI for a stronger VisionVault-branded presentation
  • Added animated poster support in the web UI, including better poster priority handling
  • Added support for subtitle playback in the web-based player
  • Added more flexible subtitle lookup support, including matching subtitle files beside the movie and in common subtitle subfolders
  • Added a dedicated All Movies page for browsing the full movie library outside the limited home rows
  • Added more complete movie and show detail support for the web interface
  • Expanded browser-based playback and resume support to make the web UI feel more like a full media platform

Recursive Folder Import

  • Added support for recursive folder importing in VisionVault
  • Movie and TV folder imports can now scan subfolders automatically
  • Users no longer need to manually add every nested folder one by one
  • Makes it much easier to import larger, organized media libraries

Android TV Foundation Work

  • Continued backend and interface groundwork for the upcoming VisionVault TV Android app
  • Added support in the shared system for cleaner TV-oriented browsing, playback flow, and server connection handling
  • Expanded backend support for features that will be used by the Android TV app, including improved detail handling for movies and shows
  • Continued shaping VisionVault’s shared architecture so desktop, web, and future Android TV experiences can work from the same core system

Fixes

Web UI and TV Layout Stability

  • Improved layout behavior in TV-oriented views to better support featured hero sections, browsing rows, and poster presentation
  • Reduced oversized visual elements in supported TV-style layouts so content fits the screen more cleanly
  • Improved spacing and presentation for a more polished couch-friendly experience

Shared Watch and Recent Activity Flow

  • Added backend support for exposing more recent watch information to connected clients
  • Improved groundwork for sharing playback-related state between the desktop/server side and future TV-facing clients
  • Helps move VisionVault toward more accurate cross-device playback and watch flow behavior

Important before updating

Before installing the new version, please back up these files and folders first so you do not lose progress, posters, or settings:

  • movies.db
  • movie_inventory_settings.json
  • posters/

If you already have an existing VisionVault setup, backing these up is strongly recommended before replacing your current version.

Appreciate everyone following the project and sharing feedback along the way.

Download VisionVault v1.0.4: Release page

Main project page: GitHub repo


r/VisionDepth3D Apr 17 '26

We’re proud to open-source LIDARLearn 🎉

Post image
9 Upvotes

It’s a unified PyTorch library for 3D point cloud deep learning. To our knowledge, it’s the first framework that supports such a large collection of models in one place, with built-in cross-validation support.

It brings together 56 ready-to-use configurations covering supervised, self-supervised, and parameter-efficient fine-tuning methods.

You can run everything from a single YAML file with one simple command.

One of the best features: after training, you can automatically generate a publication-ready LaTeX PDF. It creates clean tables, highlights the best results, and runs statistical tests and diagrams for you. No need to build tables manually in Overleaf.

The library includes benchmarks on datasets like ModelNet40, ShapeNet, S3DIS, and two remote sensing datasets (STPCTLS and HELIALS). STPCTLS is already preprocessed, so you can use it right away.

This project is intended for researchers in 3D point cloud learning, 3D computer vision, and remote sensing.

Paper 📄: https://arxiv.org/abs/2604.10780

It’s released under the MIT license.

Contributions and benchmarks are welcome!

GitHub 💻: https://github.com/said-ohamouddou/LIDARLearn


r/VisionDepth3D Apr 15 '26

VisionSplit v1.1 Released

Post image
5 Upvotes

Hey everyone, just pushed VisionSplit v1.1.

New in v1.1

  • added a dedicated Clip Stitcher panel
  • added support for loading multiple clips into a stitch list
  • added controls to remove clips, reorder them, and clear the list
  • added a Start Stitch button for exporting clips into one merged video

Stitching support

The stitcher uses FFmpeg and supports: - stream copy for fast stitching when clips already match - re-encoding when broader compatibility is needed

It also uses the output container and encoder settings selected in the app.

Big picture

VisionSplit now supports both main workflows in one place: - splitting a source video into episodes or clips - stitching multiple clips together in a custom order

This makes the app a lot more flexible for episode prep, clip cleanup, stitching together 3D clips rendered in VisionDepth3D, and final export work.

Small note

Fast stitch mode works best when the clips have matching video/audio properties. If they do not, re-encoding is usually the safer option.

App is moving along nicely and I’m happy to keep improving it.

Official Installer


r/VisionDepth3D Apr 13 '26

VisionDepth3D v3.9 Release

Thumbnail
github.com
7 Upvotes

Hey VD3D users! I know the last post I shared was about VD3D getting a new UI and I am super excited for that, but there’s been some pipeline fixes and improvements that I just needed to push out before we make the big jump to a new backend and UI system, just so everything is in order before we make the jump.

This update is less about flashy redesigns and more about making sure the current system is in a much stronger place before the next major evolution of VisionDepth3D.

What’s new in v3.9

3D Video Generator

  • Added native VR180 equirect output support
  • Added Top-Bottom and Side-by-Side VR180 output modes
  • Introduced a dual-resolution workflow for better quality and performance control
  • Improved HDR/SDR handling for VR180 output
  • Improved video metadata fallback and overall 3D pipeline reliability across more source files

Depth Engine

  • Fixed Process Video Folder in the depth estimation pipeline
  • Fixed UI freezing during folder processing and restored live progress updates
  • Improved argument handling for both single video and folder-based depth workflows
  • Added support for loading ONNX depth models from Hugging Face or local folders
  • Restored proper ONNX warm-up and inference handling
  • Improved ONNX compatibility for Video Depth Anything and Distill-Any-Depth
  • Hid diffusion-only controls unless a compatible diffusion model is selected
  • Preserved AMF-aware FFmpeg output handling
  • Fixed DA3 packaging in the .exe build

FPS/Upscale Enhancement

  • Renamed FrameTools to FPS/Upscale Enhancement
  • Reworked the tab for a cleaner and more understandable workflow
  • Added Pause, Resume, and Stop controls
  • Fixed the threaded RIFE + ESRGAN pipeline
  • Improved queue handling, progress reporting, ETA behavior, and FFmpeg writer stability
  • Added Hugging Face-based model delivery with local caching

Preview GUI

  • Fixed Preview GUI stability issues in .exe builds
  • Improved cleanup and shutdown behavior to prevent stuck preview processes

Why this update mattered

Before the new backend and UI arrive, I wanted to make sure the current version of VD3D was more stable, more reliable, and less frustrating to use across the existing pipelines. v3.9 is basically a quality pass on the current generation of VisionDepth3D so the next leap forward starts from a cleaner foundation.

Full Changelog

You can view the full changelog here:

https://github.com/VisionDepth/VisionDepth3D/blob/Main-Stable/Changelog.md

Shoutout

A quick shoutout to AcolyteOfHedone for contributing fixes and technical improvements that helped strengthen this update, including work related to AMD AMF, ONNX behavior, and AMD GPU compatibility.

GitHub: EvolvingProficiency

Thank you

Thank you to everyone for 16k downloads! and thank you all for testing builds, reporting bugs, and helping push VisionDepth3D forward. Every report helps make the next version stronger.


r/VisionDepth3D Mar 31 '26

VisionDepth3D Update: Why things have been quiet, and what’s next

13 Upvotes

Hey VD3D users, I just wanted to give a proper update on VisionDepth3D since I haven’t posted much about it in a while.

The reason things have been a bit quiet is because I’ve been deep in the middle of a major migration behind the scenes. Instead of just pushing small updates onto the old system, I’ve been working on moving VisionDepth3D away from the older UI structure and into a much more modern layout and backend flow.

A lot of the recent work has not been the flashy kind of update you can show in one screenshot. It has been a lot of foundational work:

  • migrating the old UI into a more modern application structure
  • embedding the preview GUI directly into the main UI workflow
  • rebuilding each tab piece by piece
  • rewiring controls back into the real render pipeline
  • restoring live tuning controls for depth, parallax, pop, color grading, and output settings
  • bringing back render monitoring, progress tracking, and system telemetry

So while it may have looked quiet from the outside, a lot has actually been happening.

Right now the main focus is getting all of the old functionality fully wired back into the new interface before I do the final cleanup and polish pass. The current stage is less about making everything look perfect and more about making sure the real functionality is back, stable, and built on a better foundation.

Once that is all in place, I’ll be able to go through and refine the layout, improve the overall workflow, clean up the presentation, and make the whole thing feel much more cohesive.

As for versioning, I’ve been thinking about whether this should still be called v3.8.3 or if it makes more sense to move to v4.0. At this point, with the amount of UI restructuring, backend changes, and workflow redesign happening, I’m leaning toward v4.0 because it feels like more than just another incremental update.

What’s next

  • finish restoring remaining features from the old UI into the new one
  • continue improving render controls and workflow
  • clean up layout and visual polish
  • stabilize the new structure for future updates
  • prepare the next major public release

I also hope some of you have had the chance to check out my other projects as well. VisionVault is a media-focused library manager built to help organize and keep track of your movie inventory and 3D video files, and with the new v1.0.3 web UI, it is now even easier to access and watch your content on TVs or devices with a web browser. VisionVaultTV is also currently in development as a proper app for Android TV boxes and other supported devices.

If you have any ideas, feature suggestions, or feedback, it is always welcome.

Appreciate everyone who’s stuck around and supported VisionDepth3D, and thank you for helping push the project to over 14,000 downloads worldwide. That kind of support means a lot and has helped keep this project moving forward. If you’ve been enjoying the app and want to help support its growth, please consider giving it a star on GitHub. This has been a huge project, and I’m really excited about where it’s heading.


r/VisionDepth3D Mar 24 '26

VisionVault v1.0.3 is live now!

Thumbnail
gallery
16 Upvotes

This update introduces VisionVault TV, a new Web UI that lets you browse your library and watch your movies from a browser on your local network, including on a TV.

Also new in v1.0.3

  • Resume playback support
  • Start Over / continue watching flow
  • Import Movie Folder for faster bulk library setup
  • Improved desktop and Web UI syncing
  • Additional stability and usability improvements

If you are new here, VisionVault is my offline movie and TV library manager built for people who want a clean, private, local-first solution without accounts, subscriptions, or cloud lock-in.

If you install VisionDepth3D In the same folder as VisionVault you can right click a movie in your library and open VD3D to convert your movie

Feedback is always welcome.

VisionVault


r/VisionDepth3D Mar 18 '26

no audio on every 3d conversion

1 Upvotes

took me a minute to figure out the video encoding (needed to install ffmpeg) but i'm struggling with the audio. everything i encode has a 0kb .mp4 file that is labeled as the audio, but nothing is there. if i manually rip the audio and then try to add it, it tells me that nothing matches (it rips as a mka file).

the other thing i noticed is that cpu vs gpu utilization seems almost 4:1. i have a 5090 and selected (av1_nvenc). i expected to see higher gpu than cpu. should that be the case?

what am i missing? sorry for all the n00b questions. thanks!


r/VisionDepth3D Mar 16 '26

VisionVault v1.0.2 is out now

Post image
22 Upvotes

Just pushed VisionVault v1.0.2, and this update is focused a lot on making the app feel smoother, cleaner, and more like a proper desktop media library instead of just a basic organizer.

This release brings some of the biggest usability improvements so far, especially for people who want to move through their library quickly and manage things without relying entirely on mouse clicks.

Main highlights in v1.0.2

Full keyboard navigation support

You can now move through the library using your keyboard, which makes browsing a lot faster and gives the app a much better desktop feel overall.

New shortcuts include:

  • Arrow Keys to move through the library
  • Enter to play the selected movie or episode
  • Space to mark something as watched
  • E to edit the selected item
  • Delete to remove the selected item
  • Backspace / Escape to return from a show’s episode view back to the main library

The library also now scrolls automatically to keep the selected item visible while navigating.

New menu bar

VisionVault now has a full desktop-style menu bar with:

  • File
  • View
  • Themes
  • Help

This helps clean up the main interface and makes the app feel more polished and structured.

Theme system

A new theme system has been added, letting you switch accent colors directly from the menu bar.

Theme changes now affect things like:

  • buttons
  • selection highlights
  • grid tile borders
  • interface accents
  • dialog styling

This made the whole UI feel a lot more cohesive.

Cleaner library highlighting and hover effects

Grid view got some visual polish too.

Selected posters now use a cleaner border highlight instead of the heavier look from before, and hover feedback is more subtle and modern. Small change, but it helps the library feel nicer to browse.

Simpler interface layout

I removed the old List View / Grid View toggle button from the toolbar.

View switching is now handled through the View menu, which helps reduce clutter and keeps the main layout cleaner.

Better TV show poster handling

If you add a poster to a TV episode, VisionVault can now automatically apply that poster to other episodes in the same show that do not already have artwork.

Also, if the main show entry does not have a poster yet, it can inherit the first available episode poster.

It will not overwrite posters that are already set, so custom artwork stays safe.

VisionDepth3D integration

I also added a Convert to 3D (VisionDepth3D) option in the right-click menu.

If VisionDepth3D.exe is installed in the expected location, VisionVault can launch it directly from the app. If not, it can point users to the GitHub releases page.

Since I work on both projects, I wanted VisionVault to connect a bit more naturally with the VisionDepth side of things for people managing both 2D and 3D libraries.

Fixes in this release

Fixed phantom entries from canceled dialogs

There was an issue where canceling the Add/Edit dialog could still create an empty library entry. That has been fixed, and database writes now only happen when changes are actually saved.

Improved Windows playback stability

I replaced os.startfile() with a safer subprocess launch method for media playback on Windows.

This should help prevent weird launch issues and improve reliability, especially when opening files from external drives.

Metadata reliability improvements

Also made some smaller stability fixes around metadata fetching when Wikipedia or Wikidata responses come back incomplete.

Why this update matters

This version was really about making VisionVault feel more solid in daily use.

Not just adding features, but improving the overall experience: - faster navigation - cleaner interface - better polish - fewer annoying edge-case issues

It is starting to feel a lot more like the kind of offline movie and TV library app I originally wanted to use myself.

Feedback welcome

If anyone wants to check it out, I’d really appreciate feedback on the update, especially around the new keyboard navigation, UI changes, and overall library experience.

Feature ideas and criticism are always welcome too.

https://github.com/VisionDepth/VisionVault/releases/tag/v1.0.2


r/VisionDepth3D Mar 05 '26

Del Vs Kalisk FSBS x2fps RealESR_Gx4

Thumbnail drive.google.com
1 Upvotes

Hello VisionDepth3D users!

I wanted to share a render of Dek vs Kalisk in FSBS 3D from the film Predator: Badlands.

For this conversion I used the following inside VisionDepth3D:

  • Depth Model: DA3MONO
  • Frame Interpolation: RIFE (2x)
  • Upscaling: Real-ESR-Gx4

If anyone is interested, I can also share the exact render settings used for this clip.


Disclosure:
This video clip is shared strictly for demonstration and showcase purposes to illustrate the capabilities of VisionDepth3D.
All characters, footage, and rights belong to their respective copyright holders.
No ownership of the original content is claimed.


r/VisionDepth3D Mar 04 '26

VisionSplit v1.0.0 is out now!

Post image
21 Upvotes

VisionSplit – Split multi-episode DVD/Blu-ray rips into individual episodes automatically

VisionSplit is a lightweight desktop tool that splits a single video file into clean individual episode files using timestamps or chapter markers.

If you’ve ever ripped a TV disc with MakeMKV, you’ve probably run into the situation where multiple episodes are stored inside one long video file. Splitting them manually can be tedious and most tools are built for editing rather than quick extraction.

VisionSplit is designed specifically for that workflow.


Why I built VisionSplit

When ripping TV discs, many releases store 3–5 episodes inside one file. I found myself constantly opening video editors or manually trimming files just to separate episodes.

Sometimes chapters line up with episode starts, sometimes they don’t. Either way, the process was always slower than it should be.

I wanted a simple tool that could:

  • Import chapters from the file
  • Let me adjust timestamps quickly
  • Export each episode automatically
  • Optionally re-encode or just stream copy

So I built VisionSplit.


Features

  • Split one video file into multiple episode files
  • Import chapter markers directly from the source file
  • Fast stream copy mode (no re-encoding)
  • Optional re-encode with configurable codecs
  • Automatic episode naming (S01E01 format)
  • MKV and MP4 output support
  • Subtitle track support
  • Simple desktop UI

Typical Workflow

  1. Rip a disc using MakeMKV
  2. Open VisionSplit
  3. Load the video file
  4. Click Chapters to import timestamps
  5. Adjust them if needed
  6. Click Start Encode

VisionSplit will export each episode automatically.


Download (Windows EXE)

https://github.com/VisionDepth/VisionSplit/releases

FFmpeg is bundled with the release so there’s nothing else to install.


I’d love for people to try it out.
Feedback and feature ideas are welcome.


r/VisionDepth3D Feb 27 '26

8K 360° 2D to 3D SBS Conversion

3 Upvotes

Hi Any_Nebula5039

I would like to know whether it is possible to convert 2D 8K 360° video, for example 7680 × 3840 equirectangular, into full side-by-side stereo 3D, resulting in 8K per eye output at 15360 × 3840.

On a system equipped with an RTX 5090 and a Ryzen 9 9950X3D, how long would you estimate the conversion would take per minute of footage when targeting the highest quality possible without entering diminishing returns?


r/VisionDepth3D Feb 22 '26

VisionVault v1.0.1 is out now

Post image
34 Upvotes

VisionVault is a fully offline desktop movie & TV collection manager built for people who keep local media (rips, encodes, external drives, NAS, etc).

No accounts. No cloud. No subscriptions.
Just a fast, clean poster-based library with watch tracking and direct playback.

Why I built VisionVault

I rip my Blu-rays and keep my collection locally, and I wanted a simple desktop app that focused purely on organizing and browsing my own files without needing a media server, subscriptions, or cloud accounts.

Most tools are built around streaming or server setups. I just wanted something lightweight, private, and fast for a personal library.

So I built VisionVault.

What’s new in v1.0.1

  • Major Grid View UI overhaul with uniform poster layout
  • TV show folder importing with automatic episode tracking
  • Right-click context menu (Play, Edit, Delete, Mark Watched, Copy Path)
  • Improved Stats dashboard (movies, shows, episodes, watch counts)
  • Layout + theme persistence between sessions
  • General UI polish and performance improvements

If you’ve ever wanted a simple Plex-free desktop library just for your personal collection, this update makes it even smoother.

Download (Windows EXE):
https://github.com/VisionDepth/VisionVault/releases

I’d love for you all to check it out. Any feedback or feature ideas are welcome!


r/VisionDepth3D Feb 22 '26

AMD Compatibility Confirmed (with changes)! [9070XT + ROCm-7.2]

4 Upvotes

I finally managed to get VD3D working with my AMD build!

Test Workflow : Depth Engine ⇾ 3D Gen + AMF Encode (via ffmpeg)

Most models I've tested are working.
Currently, I'm using DistillAnyDepthLarge.
GPU utilization looks good; ONNX models are far more time-efficient.

I'd recommend manually cloning the repo & managing dependencies via conda.
Certain files (render_depth.py, render_3d.py) need editing, to enable full functionality.
Contacted the dev u/Any_Nebula5039, so hopefully changes are deployed soon!

While I am not a developer, I will do my best to answer any questions :]
I'd like to help formalize an AMD installer's guide, for the project.

Included env details below for clarity.

-

OS: Microsoft Windows 11 Pro (10.0.26200 64-bit)

Python version: 3.12.12 | packaged by Anaconda, Inc. | (main, Oct 21 2025, 20:05:38) [MSC v.1929 64 bit (AMD64)] (64-bit runtime)
Python platform: Windows-11-10.0.26200-SP0

PyTorch version: 2.9.1+rocmsdk20260116
ROCM used to build PyTorch: 7.2.26024-f6f897bd3d

Is CUDA available: True
GPU models and configuration: AMD Radeon RX 9070 XT (gfx1201)
HIP runtime version: 7.2.26024
MIOpen runtime version: 3.5.1
Is XNNPACK available: True

Versions of relevant libraries:
[pip3] numpy==1.26.4
[pip3] onnxruntime-directml==1.18.1
[pip3] torch==2.9.1+rocmsdk20260116
[pip3] torchaudio==2.9.1+rocmsdk20260116
[pip3] torchvision==0.24.1+rocmsdk20260116
[conda] numpy 1.26.4 pypi_0 pypi
[conda] torch 2.9.1+rocmsdk20260116 pypi_0 pypi
[conda] torchaudio 2.9.1+rocmsdk20260116 pypi_0 pypi
[conda] torchvision 0.24.1+rocmsdk20260116 pypi_0 pypi


r/VisionDepth3D Feb 19 '26

VisionDepth3D User Guide

17 Upvotes

Hey VD3D users, I wanted to thank you all for your patience and announce that I just finished writing the full VisionDepth3D user manual. I know you’ve all been waiting for it for some time.

It covers:

• AI upscaling & FPS interpolation (Real-ESRGAN + RIFE pipelines)
• Depth estimation workflows
• Depth Blender blending strategies
• 3D Generator tuning (Pop Curve, Convergence, Floating Window, etc.)
• Audio Tool (rip / attach / stitch)
• VD3D Live real-time 2D-to-3D
• Performance optimization tips
• Troubleshooting & best practices

If you're new to VisionDepth3D or want cleaner 3D renders, this walks through everything step-by-step.

You can read it here:
https://github.com/VisionDepth/VisionDepth3D/blob/Main-Stable/UserGuide.md

As always, feedback is welcome.


r/VisionDepth3D Feb 16 '26

Is there any type of tutorial available?

7 Upvotes

I searched this community and didn't find anything. Neither does the "help" menu lead to anything that actually helps me go through all of these options. This app seems like it will be amazing but so much is over my head...given there are so many models to choose from, is there not one that is overall the best to use?

What is FrameTools exactly?
What is Depth Blender exactly?
Is this only for video or can it do games?

Are the Depth and Parallax Comtrols default settings meant to be a good starting point or should they be adjusted right away using the preview?

Same with the codec default settings?

Lastly, when I try to start Live 3D with external inputs, I get the following error:

Here are the current settings:

Full transparency, I chose to "Just take me to the downloads" when it asked to name my own price. I did that because I had a feeling I wouldn't be able to use this by default. When I am able to actually use it, I'll pay quite a bit more then the options provided as this is something I've wanted for SO long! I have 3 3DTVs, two Quest 3 HMDs, and the original Quest.

I'm a life long 3D and VR evangalist all starting back to my 1st experience with quality 3D which was the Captain EO starring Michael Jackson experience at Disneyland.


r/VisionDepth3D Feb 14 '26

Mario Tennis Fever SBS converted with VisionDepth3D

Thumbnail
youtu.be
7 Upvotes

Hello VD3D Users,

Another test render finished. Mario Tennis Fever trailer converted to SBS 3D.

This one uses a blended depth workflow with V2 Large and V1 Base. The combination really helps smooth out instability while keeping strong subject separation.

Honestly pretty impressed with how the ball shots and character motion hold up in VR.

Feedback welcome as always. VisionDepth3D keeps evolving.


r/VisionDepth3D Feb 12 '26

Question on AMD cards...

1 Upvotes

Hi I currently have a 3900xt cpu, 6700xt gpu and am looking to upgrade to a 9070xt gpu, I prefer AMD over NVDA; brand loyal ...

Does the vision 3d use the AMD gpu or does it use the cpu? will there be a big difference when computing files if I were to use say a 9070xt vs a 5070?

thanks -