Antonio Sanchez
|
952eda443b
|
Fix GPU build failures.
|
2025-03-09 17:04:41 -07:00 |
|
Antonio Sanchez
|
d2ce4faa5a
|
Fix cuda 9+ builds
Fix removed `shfl_` intrinsics, disable warnings, update CUDA header inclusion.
|
2025-03-02 16:21:48 -08:00 |
|
Antonio Sanchez
|
23b1682723
|
Fix cuda device warnings
|
2025-02-28 22:09:30 -08:00 |
|
C. Antonio Sanchez
|
109935bfce
|
Fix Tensor docs
(cherry picked from commit 42d9cc0b1d)
|
2025-02-25 21:21:49 -08:00 |
|
Antonio Sanchez
|
339d7188ed
|
Fix up all doxygen warnings.
|
2025-02-25 21:05:40 -08:00 |
|
Christoph Hertzberg
|
71d0402e3e
|
Avoid throwing in destructors (this caused build warnings in test-suite)
|
2019-06-28 11:55:38 +02:00 |
|
Christoph Hertzberg
|
6870a39feb
|
Hide some annoying unused variable warnings in g++8.1
(grafted from a7779a9b42
)
|
2019-01-29 16:48:21 +01:00 |
|
Christoph Hertzberg
|
d107a371c6
|
Fix most Doxygen warnings. Also add links to stable documentation from unsupported modules (by using the corresponding Doxytags file).
|
2018-10-19 21:10:28 +02:00 |
|
Christoph Hertzberg
|
fcc41f1b9a
|
Fix a lot of Doxygen warnings in Tensor module
(grafted from 3f2c8b7ff0
)
|
2018-10-09 20:22:47 +02:00 |
|
Christoph Hertzberg
|
3b92f547f5
|
Fix more shadowing typedefs
|
2018-09-08 23:47:53 +02:00 |
|
Christoph Hertzberg
|
5be00b0e29
|
Product of empty array must be 1 and not 0.
|
2018-08-30 17:14:52 +02:00 |
|
Christoph Hertzberg
|
03326d9155
|
Fix integer conversion warning
|
2018-08-30 17:12:53 +02:00 |
|
Rasmus Munk Larsen
|
fea50d40ea
|
Fix oversharding bug in parallelFor.
(grafted from 5418154a45
)
|
2018-06-20 17:51:48 -07:00 |
|
Benoit Steiner
|
a7144f8d6a
|
Made the TensorStorage class compile with clang 3.9
(grafted from de7b0fdea9
)
|
2017-02-28 13:52:22 -08:00 |
|
Gael Guennebaud
|
106ba41c2a
|
Fix typo.
(grafted from 478a9f53be
)
|
2017-02-28 09:32:45 +01:00 |
|
Gael Guennebaud
|
d367ecb475
|
Silent warning.
(grafted from a811a04696
)
|
2017-02-20 10:14:21 +01:00 |
|
Gael Guennebaud
|
f9d655a8c8
|
Fix compilation.
(grafted from f8a55cc062
)
|
2017-02-18 10:08:13 +01:00 |
|
Benoit Steiner
|
dcc14bee64
|
Fixed the formatting of the code
|
2016-11-08 14:24:46 -08:00 |
|
Luke Iwanski
|
912cb3d660
|
#if EIGEN_EXCEPTION -> #ifdef EIGEN_EXCEPTIONS.
|
2016-11-08 22:01:14 +00:00 |
|
Luke Iwanski
|
1b345b0895
|
Fix for SYCL queue initialisation.
|
2016-11-08 21:56:31 +00:00 |
|
Luke Iwanski
|
1b95717358
|
Use try/catch only when exceptions are enabled.
|
2016-11-08 21:08:53 +00:00 |
|
Mehdi Goli
|
d57430dd73
|
Converting all sycl buffers to uninitialised device only buffers; adding memcpyHostToDevice and memcpyDeviceToHost on syclDevice; modifying all examples to obey the new rules; moving sycl queue creating to the device based on Benoit suggestion; removing the sycl specefic condition for returning m_result in TensorReduction.h according to Benoit suggestion.
|
2016-11-08 17:08:02 +00:00 |
|
Mehdi Goli
|
0ebe3808ca
|
Removed the sycl include from Eigen/Core and moved it to Unsupported/Eigen/CXX11/Tensor; added TensorReduction for sycl (full reduction and partial reduction); added TensorReduction test case for sycl (full reduction and partial reduction); fixed the tile size on TensorSyclRun.h based on the device max work group size;
|
2016-11-04 18:18:19 +00:00 |
|
Benoit Steiner
|
0585b2965d
|
Disable vectorization on device only when compiling for sycl
|
2016-11-02 11:44:27 -07:00 |
|
Mehdi Goli
|
51af6ae971
|
Fixed the ambiguity in callig make_tuple for sycl backend.
|
2016-10-31 16:35:51 +00:00 |
|
Benoit Steiner
|
0a9ad6fc72
|
Worked around Visual Studio compilation errors
|
2016-10-28 07:54:27 -07:00 |
|
Benoit Steiner
|
b0c5bfdf78
|
Added missing template parameters
|
2016-10-28 03:43:41 +00:00 |
|
Gael Guennebaud
|
530f20c21a
|
Workaround MSVC issue.
|
2016-10-27 21:51:37 +02:00 |
|
Benoit Steiner
|
0a4c4d40b4
|
Removed a template parameter for fixed sized tensors
|
2016-10-26 18:47:37 -07:00 |
|
Benoit Steiner
|
5f2dd503ff
|
Replaced tabs with spaces
|
2016-10-25 20:40:58 -07:00 |
|
Benoit Steiner
|
1644bafe29
|
Code cleanup
|
2016-10-25 20:36:14 -07:00 |
|
Benoit Steiner
|
cf20b30d65
|
Merge latest updates from trunk
|
2016-10-20 09:42:05 -07:00 |
|
Benoit Steiner
|
d3943cd50c
|
Fixed a few typos in the ternary tensor expressions types
|
2016-10-19 12:56:12 -07:00 |
|
Mehdi Goli
|
e36cb91c99
|
Fixing the code indentation in the TensorReduction.h file.
|
2016-10-14 18:03:00 +01:00 |
|
Luke Iwanski
|
e742da8b28
|
Merged ComputeCpp into default.
|
2016-10-14 13:36:51 +01:00 |
|
Mehdi Goli
|
524fa4c46f
|
Reducing the code by generalising sycl backend functions/structs.
|
2016-10-14 12:09:55 +01:00 |
|
Benoit Steiner
|
7e4a6754b2
|
Merged eigen/eigen into default
|
2016-10-12 22:42:33 -07:00 |
|
Benoit Steiner
|
5c68051cd7
|
Merge the content of the ComputeCpp branch into the default branch
|
2016-10-07 11:04:16 -07:00 |
|
RJ Ryan
|
e2e9cdd169
|
Fully support complex types in SumReducer and MeanReducer when building for CUDA by using scalar_sum_op and scalar_product_op instead of operator+ and operator*.
|
2016-10-06 10:49:48 -07:00 |
|
Benoit Steiner
|
ae1385c7e4
|
Pull the latest updates from trunk
|
2016-10-05 14:54:36 -07:00 |
|
Benoit Steiner
|
c84084c0c0
|
Fixed compilation warning
|
2016-10-05 14:15:41 -07:00 |
|
Benoit Steiner
|
8b69d5d730
|
::rand() returns a signed integer on win32
|
2016-10-05 08:55:02 -07:00 |
|
Benoit Steiner
|
ed7a220b04
|
Fixed a typo that impacts windows builds
|
2016-10-05 08:51:31 -07:00 |
|
Benoit Steiner
|
ceee1c008b
|
Silenced compilation warning
|
2016-10-04 18:47:53 -07:00 |
|
Benoit Steiner
|
6af5ac7e27
|
Cleanup the cuda executor code.
|
2016-10-04 08:52:13 -07:00 |
|
Benoit Steiner
|
2f6d1607c8
|
Cleaned up the random number generation code.
|
2016-10-04 08:38:23 -07:00 |
|
Benoit Steiner
|
2bda1b0d93
|
Updated the tensor sum and mean reducer to enable them to process complex numbers on cuda gpus.
|
2016-09-28 17:08:41 -07:00 |
|
Mehdi Goli
|
dd602e62c8
|
Converting alias template to nested struct in order to be compatible with CXX-03
|
2016-09-27 16:21:19 +01:00 |
|
Benoit Steiner
|
6565f8d60f
|
Made the initialization of a CUDA device thread safe.
|
2016-09-26 11:00:32 -07:00 |
|
Benoit Steiner
|
f6ac51a054
|
Made TensorEvalTo compatible with c++0x again.
|
2016-09-23 16:45:17 -07:00 |
|