Speech recognition jobs
## Self-Hosted Facial Recognition API – Python / PHP We need a developer to build a **self-hosted facial recognition system** using available open-source/pre-trained libraries. We do **not** want to use paid services such as Azure Face API, AWS Rekognition, Face++ etc. The system should run on our own server and provide a simple **web service/API** which our existing PHP application can call. Main requirements: * Enrol a person using one or more face images * Preferably work even when only **one enrolment image** is available * Recognise an enrolled person from a submitted image * Return Person ID / match result / confidence score * Ideally support multiple faces in one image * Store facial embeddings/templates efficiently * Fast response and suitable for many enro...
...1. WhatsApp Procurement Requests Authorized users should submit purchase requirements using: • Text messages • Voice notes • Images Each mobile number must link to a registered user, branch and business unit. 2. AI Request Extraction The system should extract: • Item • Quantity • Unit of measure • Delivery date where provided • Business unit • Requester • Instructions Voice notes require speech-to-text processing. Images require OCR / AI document understanding. 3. NetSuite Item Matching Extracted items must match against the Oracle NetSuite Item Master. Matching should support: • Exact item descriptions • Partial descriptions • Abbreviations • Spelling errors • Common kitchen terminology •...
... For example: If I say: "McDonald's didn't make money only by selling burgers." Show: Burger → restaurant → real estate/location → money/business graphic. The viewer should understand the story even better by watching the visuals. --- ## 6. TEXT / CAPTIONS Use captions throughout the video. Captions should be: - Clean - Bold - Easy to read - Properly timed with my speech - Not too small - Not covering my face Highlight important words or phrases. Example: "THIS ONE DECISION" "CHANGED THE COMPANY FOREVER" Don't make every single word a different animation. Keep the typography premium and consistent. --- ## 7. IMPORTANT WORD EMPHASIS When I say an important word, v...
I need a business-grade projector purpose-built for classrooms that lets teachers and students interact directly with the projected image. The core...on optional extras—wireless connectivity or even a 4 K optical block—so feel free to flag those if they strengthen the overall design. I’ll review progress in logical milestones: concept, detailed design, prototype, final refinements. Once each stage meets the agreed acceptance tests (touch accuracy within ±2 mm across the screen, minimum 3 000 lumens real-world output, and plug-and-play USB recognition on the three major OSs), I’ll release the sign-off so we can move to the next phase. If you have a track record in projector optics, interactive whiteboards, or similar embedded systems, let’s get ...
...& iOS): Vertical Swipe UI: Inshorts-style full-screen card layout with high-speed smooth swiping. Content Structure: Headline, 60–80 word summary, cover image/video, category tag, and source link ("Read More" via In-App Webview). Multi-Language Support: Seamless support for regional fonts (specifically Telugu Unicode fonts without rendering issues) and language switcher. Audio News (Text-to-Speech): Integrated TTS button to listen to short summaries. Interactive Modules: Live Opinion Polls with real-time percentage results. Like, Bookmark, and 1-Click WhatsApp Card Share (generating an image card with app branding & link). Notifications: Push notifications with deep-linking (OneSignal or Firebase Cloud Messaging) for breaking news. Offline Caching...
We are seeking native and fluent Marathi speakers ages 18-60 years old to participate in a speech data collection project focused on high-quality audio recordings for language and AI research purposes. The project involves recording 220 scripted utterances through a mobile-based recording application while following strict quality and recording guidelines. The pay rate for this task is $4-$5 Please complete this pre-vetting Google Form: Participants will read and record client-provided phrases in controlled acoustic environments with minimal background noise. Recordings must meet quality standards, including proper silence padding, clear audio, and no interruptions or distortions. Each participant is expected to complete the required number of recordings
...underlying logic to be flexible enough to extend to other ad types down the road. Platform • The first release must work reliably on Android. If it also happens to skip any other app ads without extra effort, that’s a bonus, but Android-wide coverage is the priority. Ad-detection approach I’m open to whatever technique you believe will be the most accurate and light on resources—image recognition, pattern recognition, accessibility-service hooks, or a smart hybrid. Explain briefly in your proposal why your chosen method is the better fit Must be able to use discord to call/voice share screen must speek good english Deliverables • An installable APK (or Android module) that silently watches the screen and taps the visible skip/close cont...
...provide documentation showing which blendshape(s) should be driven for each viseme. The final result must produce convincing speech, not simply open and close the jaw. ## 5. Eyes Please carefully check: * eye rotation * blinking * eyelid deformation * eyelids following the eyeball correctly * no eyeball clipping * no visible gaps around the eyes * natural neutral eye position The eyes are particularly important because errors here immediately make a digital human look unnatural. ## 6. Mouth / Teeth / Inner Mouth Please check and correct: * lips * teeth * tongue if available * mouth interior * jaw opening * mouth closure * lip intersection * teeth clipping Speech must look natural from both frontal and 3/4 views. ## 7. Full Body Rig The body must work reliably wi...
I have an ongoing Marathi television comedy series that needs crisp, well-timed English subtitle...deliver a clean SRT (or compatible) file that can be dropped straight onto the edit timeline. Standard subtitling tools such as Aegisub or Subtitle Edit are fine—use whatever lets you guarantee frame-accurate sync and consistent line breaks. Acceptance criteria: • Translation reflects both the literal content and the cultural flavour of the jokes. • Timing is accurate to within ±0.5 sec of speech. • File passes a spell-check and reads smoothly when played back with the episode. If this first batch goes well, the full season will follow, so consistency and reliability matter. Let me know your turnaround time per 25-minute episode and any previous co...
...Here is what the job involves: • Isolate and denoise the whispered portions without introducing artifacts or making the speech sound unnatural. • Transcribe the cleaned English dialogue with correct punctuation and speaker labels. • Flag any unintelligible words with time-stamps so I can review them quickly. I typically work with WAV and MP3 files; if you prefer a different format, let me know before we start so I can convert them for you. You’re free to use whichever tools you trust—iZotope RX, Adobe Audition, Audacity, or a machine-learning workflow such as Whisper or DeepSpeech—so long as the final audio and text meet these two acceptance criteria: 1. Whispered speech is audible at a normal listening level with background noise kep...
...heading, bullet, and line break should mirror the original sheets, but feel free to apply consistent, professional formatting (fonts, margins, spacing) so the final file is easy to read and ready for further editing. Accuracy is my top priority. Please proof-read as you type so that spelling and punctuation are correct throughout. I am not looking for OCR; manual re-typing ensures we avoid recognition errors. When you reply, highlight your experience with similar transcription or data-entry work. A brief note on the types of documents you’ve handled and any quality-control steps you follow will help me choose quickly. Deliverable: • One Word (.docx) file that faithfully reproduces the content of all pages I’ll supply high-resolution scans as soon as we agree ...
I’m sitting on a library of fresh, well-researched finance blog posts and hot-topic articles, and I want them to start earning their keep. Your task is straightforward: place natural, contextual citations of my content on reputable sites so my brand gets talked about more often and in the right circles.... steering clear of spammy syndication networks. Deliverables I need from you – A brief placement plan outlining target publications or blogs and why they make sense. – A final report listing every live URL, the exact anchor text, and the page authority metrics. – Replacement or correction if a link is removed within 30 days of delivery. If you’ve shepherded finance brands to wider recognition through citation insertion before and have editor ...
...owners, entrepreneurs and companies that may need a new website, website redesign, eCommerce solution, SEO or other digital services. Campaign approach We are looking for someone who can recommend and manage the right advertising strategy rather than simply boosting posts. Our initial approach could include: 1. Awareness / Video Campaigns Introduce Evenzia to potential customers and build recognition in the market. Example: 3 reasons your website isn't generating enough enquiries or What makes a business website actually convert? The objective is to make potential clients aware of the problems we can solve. 2. Traffic / Engagement Campaigns Once people are familiar with the brand, drive interested prospects to relevant Evenzia service or landing pages. Example: Yo...
...wrapped/PWA-based build) and note any impact on cost or timeline. • Please include Play Store submission, review, and compliance requirements in your quote — including Google Play's Data Safety declaration, which will need to accurately reflect the sensitive personal data this platform handles. • Backend integrates with a third-party large language model (LLM) API for the core analysis, plus a separate speech-to-text service for voice input. • Data protection is a hard requirement throughout — the product handles sensitive personal data and must be built with GDPR compliance in mind from the ground up (not retrofitted). • Please propose the tech stack you would use or continue with, and flag anywhere you'd want to review our current imple...
I’m setting up an Automatic Number Plate Recognition (ANPR) solution from scratch and need a developer who can take the idea through to a dependable, production-ready system. Your job will cover the full pipeline—camera input, number-plate detection, character recognition, and a clean API or dashboard where I can query, export, and integrate results. I have not locked down the final feature set yet, so I’m open to your recommendations on whether to prioritise real-time alerts, historical data analytics, or seamless links to third-party platforms. What matters most is accuracy across various lighting and weather conditions and a design that lets me scale from a single roadside camera to a multi-site installation later on. If you have previous deployments on...
...badge design in every future episode. 3.4 Presenter shot • Vertical 9:16, chest-up framing, direct eye contact with lens. • Simple, high-contrast background (plain wall or subtly branded backdrop). • Front/slightly-above soft light, no harsh shadows. 3.5 Subtitles • Bold sans font (e.g. Poppins / Mukta), white text with black stroke for contrast on any background. • Must be tightly synced to speech — no lag. 3.6 Music & SFX • Low tension/percussion loop under the talking segment — build energy, stay under the voice. • Short “ding/pop” SFX on every text pop-in. • Sting/riser just before “Final winner”, then a half-second of silence — dramatic beat. 3.7 Citation template (set up now, reuse...
I have around 200 of Spanish subtitle files that accompany corporate technical-skills training videos, and I need them polished before release. The language already sits at a solid draft level, but I want every line to follow Chicago Manual of Style conventions, read naturally to a professional audience, and stay perfectly in sync with on-screen speech. You will receive the original subtitle files (SRT) and any relevant reference material. I expect you to correct grammar, punctuation, capitalization, and terminology, flag or fix timing issues if you notice them, and preserve all formatting tags. Deliverables • A clean, final Spanish subtitle file in the original format • A marked-up version (or change log) showing every edit for easy review Turnaround and accuracy...
My Windows-based weighing application already stores a JPEG snapshot of every vehicle that drives onto t...script—provided it runs reliably on Windows. OpenCV, Tesseract, EasyOCR, YOLO, or similar toolkits are fine as long as they give me accurate results on Indian plates under varied lighting. Acceptance will be based on: • Minimum 90 % recognition accuracy on a sample set of 500 real-world Indian plate images I will supply. • API or command-line call that receives the original JPEG path and returns the detected number in text. • Clear setup instructions and commented source so my team can rebuild or fine-tune later. If you have already delivered number-plate recognition in an Indian context, especially for weighbridge or toll systems, your experie...
...badge design in every future episode. 3.4 Presenter shot • Vertical 9:16, chest-up framing, direct eye contact with lens. • Simple, high-contrast background (plain wall or subtly branded backdrop). • Front/slightly-above soft light, no harsh shadows. 3.5 Subtitles • Bold sans font (e.g. Poppins / Mukta), white text with black stroke for contrast on any background. • Must be tightly synced to speech — no lag. 3.6 Music & SFX • Low tension/percussion loop under the talking segment — build energy, stay under the voice. • Short “ding/pop” SFX on every text pop-in. • Sting/riser just before “Final winner”, then a half-second of silence — dramatic beat. 3.7 Citation template (set up now, reuse...
...user enters a topic, keyword, DOI or even an abstract snippet and instantly receives the most relevant peer-reviewed papers, neatly ranked with key metadata and brief AI-generated summaries. To make that happen, the tool should draw from open databases such as PubMed, arXiv, Crossref and any other free or licensed sources you recommend, then apply NLP techniques—semantic search, named-entity recognition, topic clustering—to surface only the most pertinent results. A clean web interface or a lightweight desktop app is fine; what matters is accuracy, speed and an intuitive workflow for researchers who will be using it daily. Key deliverables • Source-code repository with clear documentation • Working search & retrieval engine connected to major scien...
I’m building a browser-accessible tool that can take an image uploaded by the user, run it through an AI model, and immediately return classification or detection results on-screen. All core logic must live server-side, exposed through a clean REST or GraphQL endpoint, so the front-end remains lightweight and responsive across modern web browsers. Key expectations • Model accuracy matters: please start with a proven open-source architecture (e.g., YOLOv8, ResNet, EfficientDet) fine-tuned on a small sample set I’ll provide, then document how to retrain it when new data arrives. • One-click deploy: include a Dockerfile and concise README so I can spin everything up on a fresh VPS. • Results returned as JSON plus visual overlays (bounding boxes or masks) rend...
...Accuracy and consistency are critical. I will verify the JSON against the original PDFs, so the output must match every field exactly as it appears, without spelling mistakes or truncation. Feel free to employ any combination of Tesseract, ABBYY FineReader, AWS Textract, or another OCR engine that can handle mixed fonts and occasional background noise—use whatever delivers the best character recognition rate. Deliverables • A single JSON file per statement, named to match the original document • A brief summary of the processing workflow (tool versions, key settings, and any post-processing scripts) so I can reproduce the results if needed Acceptance criteria • 100 % of account-holder fields captured for each statement • No extra data (transa...
...a remote freelance opportunity where participants will engage in natural, unscripted conversations covering a variety of everyday topics, including: Travel and culture Lifestyle and daily routines Technology and innovation Entertainment and media Work and career experiences Hobbies and interests Current trends and general discussions The goal is to collect high-quality, natural conversational speech for AI and language technology development. Requirements Native German speaker from Germany Access to a good-quality recording setup Good-quality microphone required Quiet recording environment with minimal background noise Comfortable participating in natural, unscripted conversations Reliable internet connection for online recording sessions Ability to follow recording guidelines...
...a remote freelance opportunity where participants will engage in natural, unscripted conversations covering a variety of everyday topics, including: Travel and culture Lifestyle and daily routines Technology and innovation Entertainment and media Work and career experiences Hobbies and interests Current trends and general discussions The goal is to collect high-quality, natural conversational speech for AI and language technology development. Requirements Native Spanish speaker from Spain Access to a good-quality recording setup Good-quality microphone required Quiet recording environment with minimal background noise Comfortable participating in natural, unscripted conversations Reliable internet connection for online recording sessions Ability to follow recording guidelines ...
...a remote freelance opportunity where participants will engage in natural, unscripted conversations covering a variety of everyday topics, including: Travel and culture Lifestyle and daily routines Technology and innovation Entertainment and media Work and career experiences Hobbies and interests Current trends and general discussions The goal is to collect high-quality, natural conversational speech for AI and language technology development. Requirements Native German speaker from native Spanish speaker from Spain Access to a good-quality recording setup Good-quality microphone required Quiet recording environment with minimal background noise Comfortable participating in natural, unscripted conversations Reliable internet connection for online recording sessions Ability to f...
...Business API is active. I’ll need you to audit our current setup, advise on any missing Meta Business Manager steps, and connect the final number to n8n once we’re ready. Success looks like 1. A fully documented n8n workflow (export JSON) that includes trigger, NLP processing, branching logic, and CRM/webhook actions. 2. Tested live on a sandbox or approved number, showing accurate intent recognition and correct data capture for the three priority tasks. 3. Clear hand-off notes so my internal team can tweak prompts, add steps, or migrate to production with minimal guidance. If you’ve already built conversational WhatsApp automations in n8n, you’ll know the best nodes, error-handling tricks, and how to keep response times lightning-fast. That&rs...
...and language data projects. As a contributor, you’ll help collect high-quality language data by recording natural speech, participating in conversations, transcribing audio, and reviewing language content. Your contributions may support linguistic research, language preservation, and the development of language technologies. Languages Currently in Scope We are currently prioritizing speakers of: •Kurdish •Pashto •Quechua •Fulah Fluency in at least one of the languages above is required. What You’ll Do Voice Recording •Participate in natural or scripted conversations with another contributor. •Complete solo voice-recording prompts when assigned. •Record clear, natural speech using a smartphone, headset, or computer micro...
...a remote freelance opportunity where participants will engage in natural, unscripted conversations covering a variety of everyday topics, including: Travel and culture Lifestyle and daily routines Technology and innovation Entertainment and media Work and career experiences Hobbies and interests Current trends and general discussions The goal is to collect high-quality, natural conversational speech for AI and language technology development. Requirements Native German speaker from Germany Access to a good-quality recording setup Good-quality microphone required Quiet recording environment with minimal background noise Comfortable participating in natural, unscripted conversations Reliable internet connection for online recording sessions Ability to follow recording guidelines...
...Registration Consultant – 0110 Specialist Behaviour Support & 0128 Therapeutic Supports We are seeking an experienced NDIS Provider Registration and Compliance Consultant to assist our allied health organisation, Beam Therapy, with a new NDIS provider registration application. We are an established allied health organisation with an existing clinical team, including: - 2 Occupational Therapists - 1 Speech Pathologist - 2 experienced Behaviour Support Practitioners Our intended registration groups are: - **0128 – Therapeutic Supports** - **0110 – Specialist Positive Behaviour Support** As our application includes Specialist Positive Behaviour Support, we expect to undergo a **Certification Audit** and are looking for someone who can help us prepare the ...
...wordmark with a distinctive abstract symbol. The overall aesthetic should feel current, uncluttered, and versatile enough to work across digital, print, and merchandise applications. What I have in mind • Style: Modern—think flat, fresh lines rather than vintage flourishes. • Structure: A true combination of both text and icon so the elements can stand together or apart without losing brand recognition. Scope of work and deliverables 1. Three initial concepts presented in vector format (AI or SVG) with accompanying PNG previews. 2. Two rounds of revisions on the chosen direction. 3. Final hand-over of the approved logo in AI, SVG, EPS, PDF, PNG, and JPG, plus a simple usage sheet outlining clear-space, sizing, and color specs. Technical expectations Pleas...
My Windows-based weighing application already stores a JPEG snapshot of every vehicle that drives onto t...script—provided it runs reliably on Windows. OpenCV, Tesseract, EasyOCR, YOLO, or similar toolkits are fine as long as they give me accurate results on Indian plates under varied lighting. Acceptance will be based on: • Minimum 90 % recognition accuracy on a sample set of 500 real-world Indian plate images I will supply. • API or command-line call that receives the original JPEG path and returns the detected number in text. • Clear setup instructions and commented source so my team can rebuild or fine-tune later. If you have already delivered number-plate recognition in an Indian context, especially for weighbridge or toll systems, your experie...
​Smart Product Scanning (AI/OCR): ​Users can take a photo of a food item's packaging or receipt. ​The system should use OCR (Optical Character Recognition) and an LLM/Vision model to automatically extract the product name, manufacture date, and expiry date (or shelf-life duration). ​Inventory Management Dashboard: ​Categorize items (e.g., Dairy, Produce, Pantry, Meat). ​Visual indicators for item status: Fresh (Green), Expiring Soon (Yellow), and Expired (Red). ​Ability to manually add, edit, or delete items. ​AI Recipe & Usage Suggestions: ​Based on the items that are closest to expiring, the AI should suggest recipes or meal ideas to use them up. ​Notifications & Alerts: ​Push notifications or alerts sent to the user X days before an item's expiration date. ​Tech...
...comprehensive analysis of publicly available Internet content concerning a discontinued consumer product/brand that was widely marketed approximately 25–35 years ago. The objective is to determine the volume and nature of current online discussion concerning the historical brand and product. Scope of Work The research will focus on two primary questions: Extent of continuing Internet discussion and recognition. Identify and quantify publicly available Internet and social-media references concerning the historical brand/product, including consumer recollections, discussions, nostalgia, product descriptions, images, advertising references, and interest in or requests for the product. Product-category associations. Analyze the collected references to determine the product ...
I am working on sharpening my spoken German and need a native or fluent speaker to correct me in real time. My main goal is to master the trickier vowel and consonant sounds that don't exist in ...spoken German and need a native or fluent speaker to correct me in real time. My main goal is to master the trickier vowel and consonant sounds that don't exist in my native language. Either online video calls or in-person lessons, whichever you offer. During each session we may hold free conversation, or focus on a word that I have trouble with. I expect you to give feedback on clarity of my speech and my pronounciation. Please mention any preferance or restrictions for session times, duration and frequency. We can determine a time that works for us both. A teaching backgro...
...(average segment length: 5–6 seconds; maximum 8 seconds). Maintain clean silence margins around audio cuts (100ms–150ms) while avoiding speech truncation or sudden noise inclusions. Speaker & Attribute Tagging: Tag audio-level attributes (Gender: Male/Female; Role: Operator/Customer). Assign sentence-level speaker IDs (e.g., 01, C1, 02, C2 for multi-speaker dialogues). Verbatim Transcription & Special Tagging: Provide verbatim transcription matching spoken words exactly. Transcribe numbers as full words (no Arabic numerals allowed). Apply specific structural and acoustic tags: [N] for background noise [OVERLAP/][/OVERLAP] for simultaneous speech [LAUGHTER] for non-verbal laughter [PIL] for sensitive private information (addresses, phone numb...
...(average segment length: 5–6 seconds; maximum 8 seconds). Maintain clean silence margins around audio cuts (100ms–150ms) while avoiding speech truncation or sudden noise inclusions. Speaker & Attribute Tagging: Tag audio-level attributes (Gender: Male/Female; Role: Operator/Customer). Assign sentence-level speaker IDs (e.g., 01, C1, 02, C2 for multi-speaker dialogues). Verbatim Transcription & Special Tagging: Provide verbatim transcription matching spoken words exactly. Transcribe numbers as full words (no Arabic numerals allowed). Apply specific structural and acoustic tags: [N] for background noise [OVERLAP/][/OVERLAP] for simultaneous speech [LAUGHTER] for non-verbal laughter [PIL] for sensitive private information (addresses, phone numb...
I need a JavaScript module that lets customers dictate complex hardware orders directly inside my web store. Using the Gladia speech-to-text API, the module must: • stream microphone input, display a simple push-to-talk UI, and show live transcription • recognise technical terms such as lengths, diameters, part numbers, and quantities, then convert each spoken line into clean JSON that my cart can consume (e.g., { "sku": "AB-45-ST", "qty": 12, "length": "1.25m" }) • expose hooks so I can pass the JSON straight to existing checkout code • allow easy extension of the custom vocabulary so future parts or units can be added without code rewrites Deliverables 1. Well-commented source code for the Gladia in...
...operate without internet or external connections, ensuring full offline functionality. Core Requirements: Offline speech recognition & translation Embedded language models stored locally on the device. High‑quality, low‑latency translation between two languages. On‑call text display overlay Translated words appear live on the phone screen during the call. Adjustable font size and language settings. Conversation recording Record both original audio and translated transcript. Secure local storage with encryption. Export options (e.g., transcript as PDF or encrypted file). Preferred Skills: Mobile app development (Android/iOS). Experience with offline speech recognition (e.g., CMU Sphinx, Kaldi, Vosk, or similar). Knowledge of embedded transla...
...guide me on renting) a mirrorless camera, lapel or shotgun mic, basic lighting, and a compact backdrop. • Direction on set: Prompt me when needed so the Gujarati delivery stays tight and engaging. • Post-production: Trim, color-grade, clean up the audio, add captions, and weave in subtle motion graphics or stock inserts. Feel free to lean on modern AI tools such as Descript, CapCut, or Adobe’s Speech Enhance to speed things up without sacrificing quality. • Reels output: Deliver ready-to-upload vertical videos (9:16, 1080 × 1920), under 60 seconds each, in H.264 MP4 with crisp Gujarati subtitles burned in and an SRT file on the side. Acceptance criteria – Sharp 4K master plus final HD reel for each take – Captions perfectly synced...
...`` | Open wide, jaw dropped, tongue low. The most open shape. **A, I.** | | `` | Rounded and open, oval. **O.** | | `` | Small, tightly pursed, pushed forward. **U, W, Q, OO.** | | `` | Mid-open, relaxed, teeth slightly apart. Catch-all. **S, T, C, K, G, N, R.** | | `` | Closed or nearly closed, corners raised. Expression only, not a speech shape. | **Compositing** | File | State | |---|---| | `` | The character's edge light alone — the lit rim along the silhouette, on transparent. Flat single colour so it can be tinted in code. Nothing else in the file. | | `` | Soft contact shadow beneath the character, on transparent. Grey/black, no colour. | **Props** — foreground objects, same shared canvas, transparent
...showing their Google Reviews and a shot of their app Then show the video of what actually happened however it is important to ensure: 1. Baby in the video is not identifiable needs to be completely blurred 2. Voice in the video needs to be changed - so that it doesn't sound like the original person (can't sound to computerised as well, needs to sound real but just not the person talking). Add speech marks which highlights key moments i.e. "Never apologised and kept mocking the situation" "Manager admitted he said it" etc etc. • Final delivery: one master file in Full HD (1920×1080, MP4 or MOV). Does not need to be 2 minutes but the raw footage of what happened is about 90 seconds. Please share examples of similar videos you have edited ...
I have a collection of printed documents that I will...agreed-upon file (Word, Google Docs, or a plain-text editor—whatever you prefer) while preserving line breaks and basic structure so it remains easy for me to navigate later. Accuracy is critical; proofreading your own work before submission will save us both revision time. If you have experience with optical character recognition (OCR) tools such as Adobe Acrobat, ABBYY FineReader, or Tesseract, feel free to leverage them, but please make sure to correct any recognition errors. I am looking for clean, error-free text as the final deliverable. Once a batch is completed, simply return the editable files and we’ll move on to the next set. That’s all there is to it—straightforward data entry wit...
...AI Orchestration × Integration × Reliability × Security × Scalability** We want Builders who know which technologies to use, how to connect them, and how to turn them into a Digital Employee that works in a real enterprise. --- ## 8. Phase 1 Deliverables At minimum: 1. Secure integration with authorized Exchange mailboxes 2. Priority Intelligence 3. Executive Alert 4. Basic email-thread recognition and deduplication 5. VIP / critical-rule configuration 6. Organizational Intelligence 7. Weekly / Monthly Intelligence Reports 8. Multilingual email understanding 9. Feedback / correction mechanism 10. Logging and exception monitoring 11. Source code and workflow configuration 12. Deployment documentation 13. Disclosure of third-party services and recurring...
...→ Knowledge Extraction → Human Confirmation → Enterprise Knowledge Storage → Retrieval** The system should prove that information from a real enterprise meeting can become **structured, traceable, searchable, confirmable, and reusable knowledge**. --- ## 5. Capability 01 | Meeting Efficiency & Trusted Record The system should support: * Meeting audio capture and/or approved audio upload * Speech-to-Text * Speaker identification / diarization where practical * Timestamped transcript * Meeting-duration management and reminders * Executive Summary * Meeting Minutes * Action Items Time management should consider: * Planned meeting duration * Midpoint reminder * Approximately 80% time reminder * End-of-meeting reminder * Important items still lacking a c...
I’ve just completed a full brand revamp and now need an experienced Google Ads professional to drive a three-week display campaign whose sole objective is to boost brand recognition across one specific country. The launch date is locked in, so everything—strategy, set-up, optimisation, and reporting—must be ready to go the moment we flip the switch. Here’s what I need from you: • Campaign blueprint: audience research focused on geographic targeting, placement strategy, bid structure, and budget pacing for the full 21-day run. • Technical set-up: creation of display campaigns in Google Ads, conversion & impression tracking through GA4 or Tag Manager, and clear naming conventions for easy hand-off. • Hands-on optimisation: daily checks ...
My core message is simple: I only release payment when I’m genuinely happy with the work delivered. I now need a creative p...” guarantee. What matters most is a cohesive concept that makes my audience remember—and trust—this unique selling point. What I expect from you • A clear concept that ties the guarantee to an overarching brand narrative • The creative assets (copy, graphics, layouts, or code) needed to launch the concept in the agreed channels • A concise explanation of how each asset improves brand recognition and audience confidence I’ll review each milestone carefully; once I’m thrilled with the results, payment is released. If this approach fits your working style and you have fresh ideas for shining a spotlight ...
I’m assembling a corpus of real-life Malay dialogue that will be used to train AI models, and I need native Malaysian voices to create it. Your task is to simply participate in a natural-flowing conversations in Malay featuring exactly two speakers (we will coordinate you with other speaker). What I’m looking for • Authentic, unscripted conversational speech – no narration, ads, or reading aloud. • Malaysian speakers talking to each other on everyday, tourism domain. • Clean, high-quality audio with silence background. Deliverables • Completed audio files named by speaker pair and session number. • A brief metadata sheet for each file (speaker ages, genders, recording device, and a one-line topic summary). I review eve...
I have a growing brand that needs words that feel like a conversation yet still sound polished and persuasive. My main goal is to boos...messages, and any must-include details. You’ll turn that into clean, original copy with a friendly tone, free of grammar slips and fluff. First drafts should arrive in Google Docs or Word, and I expect one prompt round of revisions if needed so the final text is publication-ready. Acceptance criteria • Consistent friendly voice that aligns with the brief • Clear focus on building brand recognition (not hard-selling) • Engaging openers, solid structure, strong call-to-action where appropriate • Delivered by the agreed deadline, fully proofread and plagiarism-free If this sounds like a fit, let’s talk about...
I have a PDF that consists purely of tables—no text paragraphs or images. There are fewer than ten individual tables, and I’d like each one transferred into a clean Excel workbook using standard Excel formatting (headers in the first row, consistent column widths, and basic number/text recognition—nothing custom or complex). Accuracy is crucial; I need the data captured exactly as shown, with rows and columns preserved so the sheet is ready for basic analysis the moment I open it. Please return a single .xlsx file containing all tables on separate worksheets or in one sheet if they fit logically together; whichever approach keeps the layout clear and easy to work with.
...motivational scripts into engaging educational videos presented in a whiteboard-animation style. The end goal is simple: I paste a script, click a button, and receive a polished HD video complete with synced voice-over, on-screen hand-drawn illustrations, captions, and light background music. Scope – Build or configure the AI pipeline (Python or any reliable stack is fine) that combines text-to-speech, scene planning, whiteboard drawing generation, and video rendering. – Keep the interface lightweight—CLI, web dashboard, or even a notebook is acceptable as long as it’s repeatable in one click. – Allow basic branding options such as logo placement, color accents, and an outro slide. Key requirements • Animation style must stay true to classic w...
Top speech recognition Community Articles
30 Free Courses: Neural Networks, Machine Learning, Neural Networks, Algorithms, AI
Here is a list of free courses that will help you enhance your knowledge in data science.