AI Filmmaking
Expanding creative possibilities with AI-enhanced filmmaking.

AI as an Extension of the Creative Process
I approachAI filmmakingas an extension of the traditional creative process, combining generative technologies with more than 20 years of experience in directing, cinematography, editing and audiovisual production.
AI can help visualise ideas, generate original imagery, create performances and reconstruct scenes that would otherwise be difficult, expensive or impossible to produce. It can also support voice, audio, music, visual effects and different stages of post-production.
Each project remains human-led. The tools and models are selected according to the creative and technical needs of the work, rather than allowing a particular platform to dictate the final result.
Selected AI Filmmaking Projects
These projects demonstrate different applications of AI-enhanced filmmaking, from hybrid live-action production and continuity reconstruction to original character creation, animation, voice synthesis and sound design.
WeTransact — The Mystery of the Giant Octopus
“Johan has been taken by a huge octopus! He was just dragged down the stairs!”
A corporate action-comedy created to celebrate WeTransact’s Microsoft Partner of the Year recognition. After CEO Johan Aussenac is apparently kidnapped by a giant inflatable octopus, his colleagues launch a rescue mission through the streets of Lisbon.
Generative AI was used as a production-rescue tool to reconstruct scenes that could not be filmed, repair continuity problems and create new performances involving recognisable employees wearing complex inflatable octopus costumes.
Click to explore the full production-rescue case studyMissing scenes, continuity repair, AI performances, VFX and final post-production.
WeTransact — The Mystery of the Giant Octopusis a promotional comedy short created to celebrate WeTransact’s recognition as Microsoft Partner of the Year FY25 in the Start-Up category.
Instead of presenting the achievement through a conventional corporate announcement, the film transforms it into an absurd action-comedy. After WeTransact CEO Johan Aussenac is apparently kidnapped by a giant inflatable octopus, his colleagues launch a rescue mission through the streets of Lisbon, culminating in a final confrontation at Praça do Comércio.
The apparent monster ultimately has a less threatening purpose: ensuring that the Microsoft award reaches the people who earned it.
Production challenge:The original script was highly ambitious for a single-day shoot. The cast consisted of WeTransact employees rather than professional actors, and filming had to be coordinated around their normal work and client commitments.
The production also involved multiple Lisbon locations, live-action chase scenes and inflatable octopus costumes that depended on portable air pumps. With a limited shooting window and considerable pressure to complete the script, several important narrative moments could not be filmed, while other scenes were captured without every connecting action required for seamless continuity.
Because traditional reshoots were no longer practical, the missing material and some of the continuity problems identified during editing were resolved through a hybrid workflow combining generative AI, original live-action footage and traditional visual-effects techniques.
AI scene generation:Entirely new scenes were created to complete missing narrative moments while preserving continuity with the original cast, environments and comic tone. AI-generated versions of the performers were directed to carry out actions and deliver dialogue that had not been captured during the live-action shoot.
One particularly demanding aspect of this process was the need to recreate human performers inside inflatable octopus suits. This required maintaining consistency not only in facial identity, expression, direction of gaze and body language, but also in the design and physical behaviour of the costumes.
The shape, scale, inflation, silhouette and position of the tentacles had to remain visually coherent across different shots. The interaction between the performers’ bodies and the inflatable material also needed to feel believable so that the generated scenes could sit convincingly beside the original footage.
Continuity repair and narrative reconstruction:Because the production was following an ambitious script within a very restricted schedule, some continuity mismatches only became apparent during the editing process.
In several sequences, the live-action material did not contain all the connecting shots, character movements or visual transitions needed to maintain a clear and fluid progression. Changes in position, direction of movement, costume behaviour or character interaction risked interrupting the continuity of the scene.
AI was therefore used not simply to replace footage that had not been recorded, but also to create new transitional scenes and character moments that had never existed during the original production.
These additional shots helped correct continuity problems, bridge gaps between performances and locations, and preserve the internal logic of the narrative without requiring a traditional reshoot.
In some cases, entirely new character actions were generated specifically to establish cause and effect between two existing shots. This allowed the characters’ movements, reactions and interactions to remain visually understandable while preserving the rhythm and comic timing of the film.
Voice reconstruction and lip-sync:AI voice synthesis was used to recreate missing dialogue while preserving the tone, rhythm and accents of the original performances.
In one sequence, the performer preferred to use his own recorded voice. The dialogue was therefore captured separately during post-production and then carefully synchronised with the mouth movements of the AI-generated character.
VFX, tracking and compositing:The physical award presented another substantial visual challenge because it is made from transparent glass, with visible reflections, refractions and embedded logos.
AI-generated elements were combined with traditional rotoscoping, tracking, match moving and compositing to reconstruct the trophy, preserve its authentic branding and integrate it into shots that had not been captured during filming.
The generated octopus sequences also required additional compositing and visual supervision to match facial likeness, costume inflation, tentacle placement, scale, lighting and overall silhouette between the AI-created material and the live-action footage.
Additional generated elements were integrated directly into the filmed shots, while aerial images of Lisbon were combined with the original material to expand the visual scale of the chase.
Editing, motion graphics and sound:The final narrative was shaped through detailed human-led post-production. Dynamic editing, animated cartoon titles, custom motion graphics, music, layered sound effects and sound design were used to control the pace, strengthen the action and ensure that the comedic moments landed effectively.
Colour and visual continuity:The finished film combines footage recorded with professional cinema cameras, GoPro cameras, aerial material and AI-generated sequences. Colour correction, grading and compositing were used to match these different sources and create a coherent visual identity throughout the film.
This project demonstrates how AI-enhanced filmmaking can function not only as a creative tool, but also as a practical production-rescue and continuity solution. When real-world limitations prevented every planned scene and connecting action from being captured, generative technologies provided missing performances, transitions, environments and narrative elements.
The final result was only made possible by combining those tools with traditional directing, editing, rotoscoping, visual effects, sound design, motion graphics and colour finishing. AI supplied some of the missing pieces, but the film’s structure, continuity, timing, humour and final audiovisual language remained human-directed.
Role:Direction, editing, AI integration, generative scene development, continuity reconstruction, voice and lip-sync workflow, visual effects, tracking, match moving, rotoscoping, compositing, motion graphics, sound design, colour grading and final post-production.
Script:Julia do Prado / WeTransact
Production:Creative Visuals for WeTransact
Creative Visuals — Ideas that Talk and Walk
“Hey you, just look at this! I’m an idea, and I can talk and walk.”
An experimental hybrid short that brings a brand idea to life through a playful blue cat mascot walking through Lisbon. The character represents an abstract creative idea made visible, mobile and engaging.
Original live-action cinematography was combined with generative character development, AI voice synthesis, logo animation, sound design and traditional post-production.
Click to explore the creative and AI workflowLive-action filming, mascot development, AI voice, sound and brand integration.
Creative Visuals: Ideas that Talk and Walkis an experimental hybrid short that brings a brand idea to life through character-based storytelling.
Set on the streets of Lisbon, the piece introduces a playful blue cat mascot as a visual metaphor for creative ideas made visible, mobile and engaging.
The concept explores how a brand identity can move beyond a static logo or abstract message and become a living character inhabiting the real world. By combining local cinematography, generative character creation and post-production, the project presents an imaginative and cinematic way of expressing whatCreative Visualsrepresents.
Creative approach:The original footage was captured in Lisbon and used as the foundation for the piece. The concept, visual direction, pacing, brand framing and final structure were developed through a human-led filmmaking process, with careful attention to atmosphere, humour and visual identity.
AI workflow:Artificial intelligence was used to iteratively develop the blue cat mascot, integrate the character into the filmed environment and support the creation of additional visual elements such as the aviator goggles and other design refinements.
AI voice synthesis was also used to generate the character’s English voice performance, helping define its playful and distinctive personality.
Props, sound and finishing:The piece incorporates a rare vintageYashica-44camera as part of the narrative and visual identity.
The final result was shaped through traditional post-production, including editing, sound design, layered audio effects, logo animation and colour grading.
This short is an example of hybrid filmmaking: AI accelerated the creation and integration of character-based assets, while the creative intention, cinematic treatment, timing, soundscape and brand expression remained fully guided by human direction.
Role:Concept development, live-action filming, creative direction, AI character generation and integration, voice development, editing, sound design, logo animation, colour grading and post-production.
Ramma — O Inferno Arde em Mim
“O inferno arde em mim.”
A dark, cinematic music video exploring grief, memory and inner torment, filmed entirely at night at the historic Palace of the Marquis of Pombal in Oeiras.
The production was developed primarily through traditional cinematography, practical lighting and post-production. Generative AI was used only for the climactic chapel sequence, where the projected angel wings ignite and additional flames appear around the altar and within selected elements of the chapel.
The visual effects had to remain coherent throughout a forward tracking shot following the artist towards the altar, preserving the camera movement, changing perspective and interaction of the simulated firelight with the surrounding architecture.
Click to explore the hybrid cinematography and AI VFX workflowNight cinematography, practical projection, moving-camera AI effects, generative fire, compositing and colour finishing.
Ramma — O Inferno Arde em Mimis a dark, cinematic music video exploring grief, memory and inner torment.
The video was filmed entirely at night at the historic Palace of the Marquis of Pombal in Oeiras. Its underground spaces, formal gardens and baroque chapel were used as distinct narrative environments reflecting different emotional stages of the song.
Traditional production:The project was conceived primarily as a conventional live-action music video. The performances, locations, camera movements, lighting and visual atmosphere were created during a multi-day shoot using traditional filmmaking methods.
Low-light cinematography and controlled practical lighting were used to preserve the architecture of the palace while creating a dark, gothic visual language appropriate to the music.
The chapel challenge:The intended climax required the projected angel wings to transform into fire, with additional flames appearing around the altar and within selected architectural elements of the chapel.
Using real flames inside this protected historic space was neither safe nor permitted. The scene also involved a forward tracking shot, moving from the rear of the chapel towards the altar while following the artist. Any added visual effect therefore had to adapt continuously to the changing camera position, perspective and composition.
During filming, a 600-watt light projector with optical modifiers was used to cast static blue angel wings onto the altar. This practical projection established the composition, scale, light direction and visual foundation for the transformation developed later in post-production.
AI-assisted visual effects:Generative AI was used exclusively for this climactic sequence. The static projected wings were given fluid movement and transformed into burning wings, while additional flames were generated around the altar and within selected elements of the surrounding chapel.
Because the camera travels forward behind the artist, the generated wings and flames could not behave like a static overlay. Their position, scale, perspective and relationship with the architecture had to remain coherent as the camera moved closer to the altar.
The simulated firelight also had to interact convincingly with glass, metal, statues, walls and other architectural surfaces throughout the movement. Reflections, illumination and changes in intensity were developed to support the impression that the flames belonged within the filmed environment.
The generated material was not used as a complete replacement for the original shot. It was developed from the filmed composition and combined with the practical wing projection, the artist’s live-action performance, the existing chapel lighting and the original travelling camera movement.
Artist approval:The artist was initially opposed to the use of artificial intelligence in the music video. A test of the chapel transformation was therefore produced before the sequence was included in the final edit.
After seeing how the animated wings, generated fire and environmental light interaction could be integrated with the original cinematography, he approved the sequence for the completed music video.
Post-production and finishing:The generated elements were selected, refined and integrated through a traditional post-production workflow involving editing, visual effects, compositing and colour grading.
Particular attention was given to preserving the forward camera movement, the artist’s silhouette and performance, the chapel architecture, the changing perspective, the original lighting direction and the visual continuity between the practical and generated elements.
The result demonstrates a selective hybrid workflow in which generative AI resolves a specific physical and production limitation without replacing the live-action foundation of the project.
In this case, AI functioned as an additional visual-effects tool, allowing an otherwise impractical scene to be created inside a protected location while retaining the performance, cinematography, practical projection and moving-camera shot captured during production.
Role:Co-concept development, direction, cinematography, editing, post-production, AI-assisted visual effects, compositing and colour grading.
Concept:Ramma and Diogo Pessoa de Andrade
Production:Creative Visuals
The Dilemma — Hey Zé, Come Here
“Sometimes words are not enough. And expectations... well, those can surpass any reaction.”
An experimental AI-animated comedy short about the unrealistic expectations people sometimes project onto their partners.
What begins as a domestic disagreement becomes an absurd comic-book-style escape, using character animation, synthetic voices, editing and sound design to construct the performances and final punchline.
Click to explore the creative and AI workflowCharacter animation, synthetic voices, editing, sound design and comedic timing.
The Dilemmais an experimental AI-animated comedy short inspired by a familiar relationship dynamic: the unrealistic expectations people can project onto their partners, even when those expectations are never clearly expressed.
What begins as a typical domestic disagreement gradually turns into an absurd escape worthy of a comic-book hero. By exaggerating the breakdown in communication, the film transforms an everyday situation into a short piece of visual comedy.
Creative approach:The concept, script, characters, dialogue and directorial vision were developed through a traditional human-led creative process.
The timing of the performances, escalation of the argument and final punchline were shaped through editing, sound and careful control of rhythm.
AI workflow:AI-assisted character animation was used to bring movement, facial expression and physical performance to the original static characters.
AI voice synthesis was also used to create the dialogue performances and explore the intended comedic tone.
Editing, sound and finishing:The generated material was selected, organised and refined through traditional post-production.
Editing, English subtitling, music, sound effects and sound design were combined to build the atmosphere and ensure that the comic timing remained precise.
This project illustrates a hybrid filmmaking workflow in which artificial intelligence functions as a production tool rather than the author of the work.
AI reduced the time and resources required for character animation and voice production, while the concept, narrative decisions, pacing, emotional intention and final audiovisual construction remained human-directed.
Role:Concept, scriptwriting, character creation, creative direction, AI generation and animation, voice direction, editing, English subtitles, sound design and post-production.
Karl Marx Sells an iPhone
“Don’t buy a phone. Buy the crystallised unpaid labour of the proletariat.”
A short political satire in which Karl Marx appears to promote an iPhone through the language of exploitation, alienation and capitalist consumption.
The piece began with an AI-generated black-and-white image of Marx holding a modern smartphone. The static image was animated and combined with an AI-generated Russian voice performance.
After the AI generation stage, the final piece was constructed through human editing, voice synchronisation, sound design, audio post-production, colour work and old-film effects.
Originally created as a playful provocation for a communist friend, the experiment also demonstrates how generative tools can give movement and voice to historical figures for satire, period films, documentaries and other creative work.
Click to explore the historical-character and AI workflowHistorical-image recreation, character animation, synthetic voice, editing, sound design and vintage film finishing.
Karl Marx Sells an iPhoneis a short AI-generated satire that places one of the most influential critics of capitalism in the role of a commercial spokesperson for a modern consumer product.
Holding a smartphone towards the viewer, Marx describes the device not as a telephone, but as crystallised unpaid labour, a polished object of alienation and a form of oppression that the consumer can take home.
The humour comes from the collision between Marxist language and the visual conventions of commercial advertising. A historical figure associated with the critique of commodities is transformed into someone apparently attempting to sell one.
Origin of the project:The video was initially created as an inside joke for a communist friend with whom I regularly discuss politics.
The intention was not to produce a serious political statement, but a playful provocation based on the contradiction between Marx’s ideas and the symbolism of a premium smartphone as an object of consumption, technology and social status.
Concept and script:I wrote the original satirical text by adapting concepts associated with Marxist theory to the language of a product advertisement.
“Guys, don’t buy a phone. Buy the crystallised unpaid labour of the proletariat, a shiny piece of alienation. Take your oppression home.”
Although Karl Marx was German, the script was translated into Russian as a deliberate satirical and aesthetic choice. The language evokes the later Soviet and communist imagery commonly associated with Marx, while reinforcing the exaggerated tone of the imaginary vintage advertisement.
AI-generated source image:The project began with an AI-generated static image of Karl Marx holding a modern silver smartphone and presenting it towards the viewer.
The composition was designed like a direct-to-camera product advertisement: Marx holds the telephone in one hand and points towards it with the other, addressing the viewer as though demonstrating the qualities of the product.
Historical-character animation:The static portrait was animated using generative AI to create facial movement, speech, hand gestures and small changes in posture.
Particular attention was given to preserving the identity of the character, the proportions of the face and beard, the position of the telephone and the relationship between the pointing hand and the product.
The objective was not to create perfectly contemporary or polished movement. The slightly abrupt gestures and imperfect motion became part of the intended vintage aesthetic, resembling footage recorded with an early low-frame-rate camera.
Voice generation:An AI-generated male voice was used to perform the translated Russian script.
The delivery was directed to remain calm and discursive while retaining an ironic advertising tone. The contrast between the serious voice and the absurd sales pitch strengthens the satirical effect.
Editing, sound and finishing:After generating the animated material and voice with AI, I assembled and refined the final piece through traditional post-production.
The generated voice was synchronised with the animated image during editing, and the soundtrack was developed through sound design and audio post-production to integrate the speech convincingly with the visual rhythm of the scene.
The final sound treatment was therefore not simply the direct output of the voice-generation tool. The voice, timing and supporting sound elements were combined and adjusted during post-production.
Vintage visual treatment:After the AI-generated animation was completed, the material was edited and treated to resemble an early black-and-white television commercial or an advertisement from the first decades of filmed media.
Colour correction and contrast shaping were used to create the final monochrome appearance. Film grain, flicker, vignette, image instability, scratches and other old-film effects were then added in post-production.
The video is presented in a 16:9 aspect ratio, while the monochrome image, film grain, flicker, vignette, image instability and intentionally jerky movement create the impression of aged archival footage.
These visual imperfections help integrate the generated animation and give the piece the appearance of an archival commercial discovered from an alternative historical timeline.
The finished result therefore combines AI-generated imagery, animation and voice with human editing, voice synchronisation, sound design, audio finishing, colour work and visual post-production.
Beyond the joke:Although this particular video was created as a humorous political provocation, the underlying process has wider creative applications.
Generative animation can give movement and voice to historical figures reconstructed from photographs, portraits and other archival visual material. This can support period films, historical documentaries, museum installations, educational projects and fictional works.
The same approach can be used to visualise historical testimony, create stylised reconstructions or explore how people from the past might be represented within a contemporary audiovisual language.
Such uses require careful consideration of historical accuracy, context, transparency and the distinction between documented material and creative reconstruction.
This short demonstrates the process on a deliberately absurd scale: artificial intelligence transforms a static representation of a nineteenth-century thinker into the presenter of an imaginary vintage technology advertisement.
Role:Concept development, scriptwriting, translation workflow, AI image generation, historical-character animation, voice development, editing, voice synchronisation, sound design, audio post-production, colour grading, vintage film effects and final post-production.
Los Pura Pose — Posa para mim
“Claquete. Centro. Esquerda. Centro. Direita. Centro. Cima. Centro. Baixo.”
An experimental music video born from a real production created to record human facial expressions for the training of an artificial-intelligence system.
During the original five-day shoot, participants were filmed performing different predefined expressions while an audio guide instructed them to turn their faces towards the centre, left, right, up and down. Those functional recording commands later became the lyrics and rhythmic basis of a reggaeton-influenced track created with generative AI.
The song and music video were created in a single day. The finished video combines the original instructional recording, AI-generated music, Creative Commons stock footage and rhythm-driven human editing.
Click to explore the real-world origin and creative workflowFacial-expression capture, original field audio, AI-generated music, Creative Commons footage and rhythm-driven editing.
Los Pura Pose — Posa para mimis an experimental AI-assisted music video inspired by a real audiovisual data-capture production.
I was hired to assemble and coordinate a production team responsible for filming people performing different facial expressions. According to the information provided to us, the resulting recordings would be used to help train an artificial-intelligence system to recognise expressions and emotional states.
The original production:The recording process took place over five days inside an apartment adapted to contain four separate filming stations.
Three stations were located in different interior rooms and one was positioned outside on a balcony. Each location had a precisely defined level of illumination measured in lux at the participant’s face.
The stations covered lighting conditions ranging from approximately 10 lux in the darkest setup to more than 4,000 lux in exterior daylight. This allowed the same expressions and facial movements to be recorded under very different lighting conditions.
Participants were filmed at four predefined distances from the camera, marked physically on the floor. At each distance and lighting station, they performed eight expressions selected from a predefined group that included neutral, happy, angry, excited, shocked or fearful, affectionate, disgusted, disappointed and other emotional states.
For every expression, the participant had to turn their face through a precise sequence of directions so that it could be recorded from multiple angles.
The instructional recording:To maintain consistent timing throughout the sessions, an audio guide was played during every take.
The recording began by calling for the slate, which displayed the participant’s assigned identification number and the identification number of the expression being performed.
A timed voice then instructed the participant to move through the following sequence:
“Claquete. Centro. Esquerda. Centro. Direita. Centro. Cima. Centro. Baixo. Centro.”
The same sequence was repeated across the different expressions, camera distances and lighting stations. After five days of production, the rhythm and repetition of those commands had become extremely familiar.
From technical instruction to song:After returning home, I decided to transform the functional recording process into a playful musical experiment.
The exact directional commands used during filming became the basis of the lyrics. Rather than writing a conventional narrative song, I retained the mechanical language of the production:
“Claquete, centro, esquerda, centro, direita, centro, cima, centro, baixo.”
Generative AI was then used to transform those words into a reggaeton-influenced track. The result converts a practical set of instructions into a repetitive musical command that invites the viewer to move, pose or follow the directions.
The opening of the song includes a mix of the original voice recording used during the real filming sessions and the newly generated musical track. After this introduction, the piece transitions fully into the AI-generated music.
This connection between the original production audio and the finished song preserves the real-world origin of the project rather than merely imitating the language of a fictional photo shoot.
Music-video construction:The visual component was created using stock footage made available under Creative Commons licences.
The selected images feature different people, ages, appearances and expressions, visually echoing the diversity and facial-performance focus of the original recording project.
The footage was selected, reorganised and edited to respond to the musical instructions. Changes of expression, gaze, pose, direction and framing were synchronised with the spoken commands, percussion and rhythmic accents of the track.
The performers seen in the music video are therefore not AI-generated characters. The AI contribution is primarily musical, while the visual structure was created through the selection and transformation of existing Creative Commons footage in post-production.
Editing and rhythm:The repetition of “centre”, “left”, “right”, “up” and “down” provided a simple but strict framework for the edit.
Individual shots were timed and rearranged so that facial movements, poses and visual changes interact with the instructions in the song. The editing transforms otherwise unrelated stock footage into a coherent visual performance with its own rhythm, escalation and comic energy.
Created in one day:Both the song and the completed music video were developed in a single day as a spontaneous creative response to the production experience.
The project demonstrates how material originating in a highly controlled and technical workflow can be reinterpreted through generative music, archival selection and human editing.
Rather than using AI to reproduce the original assignment, the piece reverses its logic: a process designed to collect human expressions for an algorithm becomes the inspiration for a human-directed audiovisual work.
Role:Original production-team coordination, creative concept, lyrics, AI music development, Creative Commons footage research and selection, editing, audiovisual synchronisation and final post-production.

Tools Don’t Make the Craftsman
An old proverb says that “the cowl does not make the monk.” In the same way, access to sophisticated tools does not automatically create knowledge, judgement or craftsmanship.
A camera does not make someone a filmmaker, and neither does access to an AI model. Tools can expand what is technically possible, but meaningful work still depends on ideas, experience, visual judgement and the ability to make deliberate creative decisions.
Generating an image, a voice or a video clip is only one part of the process. The greater challenge is creating material that feels intentional, coherent and connected to the story, brand or message behind the project.
My role is to guide the process from the initial concept and visual language through generation, selection, editing, sound, colour and final delivery, so that the technology serves the project rather than defining it.
What AI Can Add to a Film or Video Production
- Concept development and visual research:exploring characters, environments, moods and visual directions before production begins.
- Pre-visualisation:creating style frames, storyboards, visual treatments and early versions of scenes or sequences.
- Generative video and imagery:producing original visual material for films, advertising, branded content, documentaries, music videos and online campaigns.
- Complex or inaccessible scenes:visualising historical settings, imaginary worlds, conceptual imagery or shots limited by budget, logistics, access, rights or available archive material.
- Hybrid live-action productions:combining traditionally filmed footage with AI-generated or AI-transformed elements.
- Continuity and production repair:creating missing connecting shots, reconstructing scenes and resolving visual or narrative continuity problems discovered during editing.
- Image transformation and extension:modifying environments, extending shots, developing transitions or creating alternative visual treatments from existing material.
- AI-assisted post-production:supporting editing, compositing, image restoration, cleanup, transcription, subtitles and versioning.
- Audio, voice and music:supporting dialogue restoration, voice generation, multilingual versions, sound design and the development of original musical ideas.
Live Action and Generative AI
AI filmmaking does not need to exist separately from traditional production. Some of the most interesting possibilities come from combining real cinematography, locations, performers and practical elements with generative imagery and AI-assisted post-production.
A production may begin with original live-action footage and use AI to extend a location, transform a visual element, create an additional shot, develop a transition or introduce imagery that could not be captured during filming.
The opposite approach is also possible: AI-generated material can provide the starting point for a sequence that is later refined through editing, compositing, sound design, colour correction and other traditional post-production techniques.
This hybrid approach makes it possible to retain the authenticity and detail of filmed material while expanding the visual possibilities available to the project.
AI-Enhanced Audio, Voice and Sound
Sound is an essential part of storytelling. Alongside generative imagery and video, AI can support different stages of audio production, from improving recorded dialogue to developing voices, music and soundscapes for a finished film.
Depending on the needs of the project, AI-assisted audio workflows may include:
- Dialogue cleanup and restoration:reducing noise, improving clarity and recovering recordings made in challenging conditions.
- Transcription and subtitles:accelerating the preparation of transcripts, captions and translated subtitle versions.
- Voice generation and multilingual adaptation:creating narration, temporary voice tracks, alternative language versions or carefully authorised voice reproductions.
- Music development:exploring musical ideas, moods and compositions created or adapted to the rhythm and emotional direction of a film.
- Sound design and atmosphere:developing ambience, textures and supporting sonic elements that strengthen the visual world of the project.
As with image generation, these tools are used selectively and remain part of a human-led creative process. Voice rights, performer consent, music usage and the intended distribution of the project are considered before the final material is produced or released.
A Filmmaker-Led AI Workflow
Creative direction
Every project begins by defining its purpose, audience, visual language and intended emotional response. AI is introduced only where it can genuinely strengthen the idea.
Choosing the right tools
I do not restrict the process to one platform or model. Different tools are tested and selected according to the required visual style, movement, consistency, resolution and type of production.
As generative technology continues to evolve, the specific tools may change. The creative objective, rather than the platform, remains at the centre of the workflow.
Generation, selection and refinement
Generating material is only one stage of the process. Outputs are reviewed, compared, selected and refined. When necessary, they are regenerated, transformed or combined with other visual elements until they support the intended direction.
Editing and post-production
The selected material is integrated into a professional post-production workflow that may include editing, compositing, colour correction, sound design, music, voice, titles, subtitles and final mastering.
Consistency and quality control
Particular attention is given to continuity, character consistency, composition, movement, lighting, rhythm and the visual details that distinguish intentional filmmaking from generic AI-generated content.
Where AI Filmmaking Can Be Useful
- Commercials and branded films
- Corporate and institutional videos
- Documentaries and historical reconstructions
- Music videos and visualisers
- Short films and fictional sequences
- Concept films and mood pieces
- Product visualisation
- Social media and online campaigns
- Pitch films, treatments and pre-visualisation
- Hybrid productions combining live action and generative imagery
- Production rescue and continuity reconstruction
- Multilingual voice and video adaptations
Responsible and Considered Use of AI
The use of AI in professional production should be deliberate and appropriate to the project. Depending on the content and intended distribution, this may involve considering copyright, source material, likeness and voice rights, privacy, client confidentiality and transparency with the audience.
When AI is used with client material, sensitive information or unpublished creative work, the workflow should also take into account how data is processed and stored by the platforms involved.
I avoid treating AI as a shortcut for replacing creative judgement or professional craftsmanship. The objective is to use these technologies in a way that protects the identity of the project and maintains the quality, credibility and human perspective expected by the client.
Related Film and Video Services
- Film Director & Filmmaker
- Cinematographer / Director of Photography
- Videographer Services
- Film Production Services in Portugal
- Advertising & Branded Content
- Films & Documentaries
- Client Reviews
Let’s explore how AI can expand the creative possibilities of your next project.
Tell me about itFilmmaker, Producer, Director & Cinematographer (DoP) / Videographer
Based in Lisbon, working across Portugal and internationally.
