14 KiB
Version 3.0
Release date: 16 July 2025
1. Major Features & Core Capabilities
1.1 AI-Powered Face Restoration (GFPGAN)
- New
AI_face_restorationClass: A new, specialized class has been implemented to handle face restoration models. This class is architected to manage the unique preprocessing and post-processing requirements of models like GFPGAN, distinct from standard upscaling models. - GFPGAN Model Integration: The GFPGAN v1.4 model has been added to the AI model repository and is now selectable from the UI. It is listed under a new
Face_restoration_models_listcategory. The main orchestrator (upscale_orchestrator) now detects when a face restoration model is selected and routes the task to the appropriateAI_face_restorationinstance. - Specialized Processing Pipeline: The new class introduces a dedicated pipeline for face enhancement. This includes resizing the input image to the model's required dimensions (e.g., 512x512 for GFPGAN), handling color channel conversions, and post-processing the output to restore the image to its original dimensions.
2. UI/UX Modernisation
2.1 Complete Thematic Redesign
- The application has undergone a significant visual overhaul with a new, professionally designed color scheme to improve aesthetics and user comfort during long sessions. The new theme provides better contrast and a more modern look.
| Element | New Value (v3.0) | Old Value (v2.2) |
|---|---|---|
| Background | #1A1A1A (Deep Black) |
#000000 (Pure Black) |
| App Name Color | #FF4444 (Bright Red) |
#FF0000 (Pure Red) |
| Widget Background | #2D2D2D (Dark Grey) |
#5A5A5A (Grey) |
| Accent/Border | #FFD700 (Gold) |
Gold & Red |
| Button Hover | #FF6666 (Light Red) |
background_color |
| Info Button | #B22222 (Dark Red) |
widget_background_color |
2.2 Enhanced Splash Screen
- Dynamic Progress Bar: The splash screen now features a
CTkProgressBarto provide visual feedback on the application's loading status, enhancing the startup experience. - Smooth Fade-Out Animation: A new
fade_outmethod using a cosine function has been implemented for a smooth, animated exit transition instead of an abrupt disappearance. - Improved Information Display: The splash screen now prominently displays the application version number.
2.3 Redesigned and Resizable Message Boxes
- The
MessageBoxclass was significantly improved to handle large blocks of text, such as detailed error messages. It now implements aCTkScrollableFrame, ensuring that content is always accessible without forcing the dialog to an unmanageable size. - The dialogs now have defined
minsizeandmaxsizeproperties for better window management.
2.4 Improved UI Readability
- The main AI model dropdown menu is now logically grouped by model type (Upscaling, Denoising, Face Restoration, Interpolation), with a
MENU_LIST_SEPARATORbetween categories. This makes it easier for users to find and select the appropriate AI model for their task.
3. Performance and Code Optimisation
3.1 Memory Optimisation with Contiguous Arrays
- Widespread use of
numpy.ascontiguousarrayhas been implemented across the codebase. This is applied during critical image handling steps inAI_upscale.preprocess_image,AI_interpolation.concatenate_images, and the newAI_face_restoration.preprocess_face_imageclass. This ensures data is aligned in memory, which can significantly speed up operations in backend libraries like OpenCV and ONNX Runtime.
3.2 Refined Data Type Handling
- The
AI_upscaleclass now explicitly ensures input images are converted tofloat32before normalization, improving precision and preventing potential data type mismatches during inference. - The
AI_face_restorationclass is configured to intelligently select betweenfloat16andfloat32based on the specific model's requirements (fp16: Truein config), further optimizing performance and VRAM usage for compatible models.
4. Codebase Health and Maintainability
4.1 Specialised Class for Face Restoration
- The logic for face restoration has been fully encapsulated within the new
AI_face_restorationclass, separating it from the general-purposeAI_upscaleclass. This object-oriented approach makes the code more modular, readable, and easier to extend with different face enhancement models in the future.
4.2 Robust BGRA to BGR Conversion
- The application now explicitly handles images with an alpha channel (4-channel BGRA) when using face restoration models. A new import for
COLOR_BGRA2BGRwas added, and it is used withinpreprocess_face_imageto convert images to the 3-channel BGR format expected by the GFPGAN model. This prevents runtime errors and ensures correct processing of PNGs or other images with transparency.
Version 2.2
Release date: 7 July 2025
1. Major Enhancements and Stability Overhaul
1.1 Comprehensive Logging System
-
Implemented a full-featured
loggingsystem that writes to files in the user'sDocumentsfolder (warlock_studio.loganderror_log.txt). This provides detailed diagnostics for debugging without relying solely on console output. -
A unified
log_and_report_errorfunction centralizes error handling, ensuring all critical issues are both logged and displayed to the user.1.2 Proactive Environment Validation
-
The application now performs pre-flight checks before processing begins to prevent common failures.
-
Includes validation for Python version, required modules (
validate_environment), FFmpeg availability, disk space, and available RAM (validate_system_requirements). -
Verifies that all input file paths exist and are accessible (
validate_file_paths) and that the output directory is writable (validate_output_path).1.3 Resilient Video Encoding Pipeline
-
The entire
video_encodingfunction was overhauled for maximum reliability. -
Codec Fallback: The system now tests for hardware codec availability (NVENC, AMF, QSV) before encoding. If a selected hardware encoder is not functional, it automatically falls back to the highly compatible
libx264software encoder. -
Robust Audio Handling: Implements a fallback chain for audio processing. It first attempts to directly copy the audio stream; if that fails, it attempts to re-encode it; if that also fails, it finalizes the video without audio, ensuring a video file is always produced.
1.4 Graceful Shutdown and Cleanup
-
Implemented
atexitandsignalhandlers to ensure that temporary files are cleaned up and child processes are terminated safely, even on unexpected exits. -
Replaced abrupt
process.kill()calls with the more gracefulprocess.terminate()to allow for cleaner process shutdown.
2. Performance and Memory Optimization
2.1 Aggressive Memory Management
-
Video frame processing (
upscale_video_frames_async) no longer holds large numbers of frames in RAM. It now writes small batches to disk and immediately calls the garbage collector (gc.collect()) to free memory, dramatically reducing the risk of crashes on long videos.2.2 Dynamic GPU VRAM Error Recovery
-
The AI orchestration logic can now recover from GPU "out of memory" errors during tiling. If an error is detected, it automatically reduces the tile resolution and retries the operation on that specific frame, preventing a total process failure.
3. Critical Bug Fixes
3.1 Resolved Video Encoding Race Condition
-
Fixed a critical bug where video encoding could start before all frame-writing threads were complete. The system now tracks all writer threads and explicitly waits for them to finish (
thread.join()) before beginning the final video encoding, preventing corrupted or incomplete videos.3.2 Corrected Persistent Stop Flag
-
The
stop_thread_flagis now reset (.clear()) at the start of each "Make Magic" execution, fixing a bug where a previously stopped job would prevent a new one from running.3.3 Eliminated Status Update Race Condition
-
Implemented a
threading.Lock(global_status_lock) to protect shared flags that update the GUI. This prevents race conditions where multiple threads could attempt to modify the status simultaneously.
4. UI / UX Refinements
4.1 Updated Splash Screen
-
Reduced splash screen duration to 10 seconds for a faster application start-up.
-
Corrected asset path to
Assets/banner.pngfor proper display.4.2 New Application Theme
| Element | New Value | Old Value (v2.1) |
|---|---|---|
| App name | #FF0000 (Red) |
#ECD125 (Gold) |
| Widget background | #5A5A5A (Grey) |
#960707 (Dark Red) |
| Accent/Border | Gold & Red | Blue & Red |
5. Codebase Health and Maintainability
5.1 Enhanced Checkpointing and Recovery
-
Added functions (
create_checkpoint,load_checkpoint) to save and resume the progress of video frame processing, allowing recovery from interruptions.5.2 Hardened Core Methods
-
Core methods in AI classes now include checks for
Noneinputs and feature default fallbacks (case _:) inmatchstatements to prevent unexpected errors with unsupported data.
Version 2.1
Release date: 23 June 2025
1. Major Enhancements and Stability Overhaul
1.1 Robust Error Handling
-
Model loading (
AI_upscale,AI_interpolation) wrapped intry…except FileNotFoundError, OSError; meaningful error messages propagate to GUI. -
extract_video_frames()validates file existence,cv2.VideoCapture.isOpened(), and frame count > 0. -
video_encoding()capturessubprocess.CalledProcessError, logsstderr, and continues with fallback strategies. -
Audio passthrough failures now trigger a silent audio‑less encode instead of total job abort.
1.2 Safe Thread and Process Management
-
Deprecated error‑raising thread stop replaced with
threading.Event(stop_thread_flag) polled at defined checkpoints.1.3 Resilient Core Processing
-
copy_file_metadata()now verifiesexiftool.exeavailability and the existence of source/target before execution.
2. UI / UX Refinements
2.1 Refined Colour Palette
| Element | New Value |
|---|---|
| App name | #ECD125 |
| Widget background | #960707 |
| Active border | Red (same as widget background) |
3. Code‑base Maintainability
3.1 Improved Code Organisation
-
File‑extension lists extracted to
filetypes.pyasSUPPORTED_IMAGE_EXTENSIONS,SUPPORTED_VIDEO_EXTENSIONS.3.2 Dependency and Initialisation
-
Added imports:
shutil.move,subprocess.CalledProcessError,threading.Event. -
Global variables initialised in
init_globals()for deterministic start‑up.
Version 2.0
Release date: 6 June 2025
1. Major Features
1.1 AI Frame Interpolation Support (AI_interpolation class)
-
Supports RIFE‑based ONNX models; generates 1 (×2), 3 (×4), or 7 (×8) intermediate frames.
-
Provides both real‑time preview and batch processing modes.
-
Integrates with
FrameSchedulerfor temporal upscaling pipelines.1.2 RIFE Models Integration
-
Added RIFE and RIFE_Lite to model repository.
-
RIFE_models_listenumerates available checkpoints;AI_models_listnow merges SRVGGNetCompact, BSRGAN, IRCNN, and RIFE families.
2. Enhancements
2.1 Visual/UI Redesign
-
Application renamed to “Warlock‑Studio” (with hyphen).
-
New dark palette (
#121212,#454242) with bright white text (#FFFFFF) and accent red (#FF0E0E).2.2 Version‑Specific User Preferences
-
User configuration stored as
Warlock-Studio_<major>.<minor>_UserPreference.jsonto avoid backward‑compatibility clashes.2.3 Modular and Scalable Layout System
-
Added GUI constants defined in
layout_constants.py(e.g.,OFFSET_Y_OPTIONS,COLUMN_1_5).2.4 Extended File‑Type Compatibility
-
Updated
SUPPORTED_FILE_EXTENSIONSandSUPPORTED_VIDEO_EXTENSIONSto include modern codecs (e.g., HEIC, AVIF, WebM).2.5 Improved GPU Execution Support
-
provider_optionsenumerates up to four DirectML devices (Auto, GPU 1 – GPU 4); selection persists across sessions.
3. Technical Refinements
3.1 Model List Structure
-
Menu drop‑downs now grouped by category separated by
MENU_LIST_SEPARATORfor readability.3.2 Advanced Interpolation Logic
-
Implements tree‑based frame generation (e.g., D→A‑B‑C) with dependency tracking to avoid redundant inference passes.
3.3 Improved Numeric Precision and Post‑Processing
-
Normalisation uses 32‑bit floats with epsilon guarding; RGBA conversion paths optimised using
numexpr.
4. UI / UX Refinements
4.1 Resizable Message Dialogs (MessageBox; Tk resizable(True, True)).
4.2 Improved Dialog Formatting – uniform spacing, font hierarchy, and default‑value display.
5. Minor Fixes
- Corrected typo
ttext_color→text_color. - Expanded inline comments and reorganised sections for clarity.
Version 1.1
Release date: 20 May 2025
1. Major Improvements
1.1 Program Start‑up Optimisation – launch time reduced via lazy module loading. 1.2 Model Loading Improvements – parallel prefetch and checksum verification. 1.3 General Performance Optimisation – core refactor, improved I/O scheduling, and smarter resource allocation.
2. Minor Fixes
- User‑interface tweaks for better accessibility (focus indicators, tab order).