CUDA Optimization for Kalman Filter in Lidar
Budget: ₹1,500 – ₹12,500 INR
I'm seeking a professional adept in CUDA optimization, particularly in task parallelism techniques, to enhance the performance of a parallel Kalman filter algorithm for lidar localization.
Key Responsibilities:
- Optimize the Kalman filter implementation using task parallelism in CUDA to ensure efficient processing and improved performance.
- Work closely with the existing algorithm to maintain its integrity while enhancing its speed and efficiency.
Ideal Skills:
- Extensive experience with CUDA programming and optimization.
- Deep understanding of the Kalman filter and its implementations.
- Proven track record in using task parallelism techniques.
- Strong problem-solving skills and attention to detail.
Your expertise will enable us to achieve real-time performance, critical for effective lidar localization.
Reference base paper:-
Published online 2024 Feb 2. doi: 10.3389/frobt.2024.1341689
Same paper implementation along with some performance improvement required for international paper publication
Deliverables:-
- Paper - 7pages (minimum)
- CUDA code for demo's on my laptop
- Results should be improved compare to base paper performance 10.3389/frobt.2024.1341689 like bar chart
- Flow chart or design flow diagram as per proposed techniques for parallel processing
Key Responsibilities:
- Optimize the Kalman filter implementation using task parallelism in CUDA to ensure efficient processing and improved performance.
- Work closely with the existing algorithm to maintain its integrity while enhancing its speed and efficiency.
Ideal Skills:
- Extensive experience with CUDA programming and optimization.
- Deep understanding of the Kalman filter and its implementations.
- Proven track record in using task parallelism techniques.
- Strong problem-solving skills and attention to detail.
Your expertise will enable us to achieve real-time performance, critical for effective lidar localization.
Reference base paper:-
Published online 2024 Feb 2. doi: 10.3389/frobt.2024.1341689
Same paper implementation along with some performance improvement required for international paper publication
Deliverables:-
- Paper - 7pages (minimum)
- CUDA code for demo's on my laptop
- Results should be improved compare to base paper performance 10.3389/frobt.2024.1341689 like bar chart
- Flow chart or design flow diagram as per proposed techniques for parallel processing