Commit Graph

1237 Commits

Author SHA1 Message Date
Philippe Tillet
2d4ddab4d0 [ir][print] improved pretty-printing of constants and instructions 2019-08-30 18:02:33 -07:00
Philippe Tillet
5db3a7adfe [python][examples] some more cleaning of dot product example 2019-08-30 17:05:03 -07:00
Philippe Tillet
7e0af2118c [codegen] worked around bug seemingly from nvptx/ptxas by simplifying multiplications by 1:
- Generated LLVM-IR looked correct
- Illegal addressing disappeared when running cuda-memcheck
- Illegal addressing disappeared when using nvptx-short-pointer
2019-08-30 16:45:14 -07:00
Philippe Tillet
141a823799 [python] refactoring in anticipation of pytorch support 2019-08-29 18:08:51 -07:00
Philippe Tillet
e3c953e79f [test] added more re-usable code in common/util.h 2019-08-28 18:06:36 -07:00
Philippe Tillet
d457482539 [codegen] fixed issue in double buffering pointer update 2019-08-28 17:50:45 -07:00
Philippe Tillet
59281f5794 [structure] better directory structure for tests 2019-08-27 20:33:38 -07:00
Philippe Tillet
37cbcfabd0 [examples] back to 96 TFLOPS on V100 2019-08-26 22:49:14 -07:00
Philippe Tillet
b4ae06a714 tracking down performance regression 2019-08-26 20:38:39 -07:00
Philippe Tillet
7cb73f66e2 testing some register gradient 2019-08-26 19:25:58 -07:00
Philippe Tillet
9ece3eccc6 some cleaning 2019-08-26 17:28:24 -07:00
Philippe Tillet
4075949f80 [python] basic tensorflow wrapper working 2019-08-26 16:53:49 -07:00
Philippe Tillet
0e0399f866 more tests 2019-08-26 11:00:00 -07:00
Philippe Tillet
321d268a4a more progress 2019-08-25 21:26:09 -07:00
Philippe Tillet
96b4d5e411 [examples] multiple transposition schemes now supported 2019-08-24 13:08:38 -07:00
Philippe Tillet
0b1c389894 [lang] changed array declarations from [{}] to [] 2019-08-23 20:34:24 -07:00
Philippe Tillet
44eb3891ae [lang] added support for restrict; added macros for attributes 2019-08-23 20:29:12 -07:00
Philippe Tillet
8c6bac49d1 [lang][codegen] added basic attribute support 2019-08-23 19:49:06 -07:00
Philippe Tillet
cb04ec0b3b some more cleaning 2019-08-23 19:22:38 -07:00
Philippe Tillet
732156b942 [general] rename *.cpp -> *.cc 2019-08-23 19:06:39 -07:00
Philippe Tillet
6158d96ff7 [general] cleaned include guards and added #pragma once 2019-08-23 18:08:05 -07:00
Philippe Tillet
606e799948 [LICENSING] updated license to incorporate credit for wgtcc 2019-08-23 17:56:30 -07:00
Philippe Tillet
a110a7e8cf [ir] changed type of tile shapes from constant_int* to int 2019-08-23 17:49:21 -07:00
Philippe Tillet
c9371c7234 [general] error messages no longer depend on a program name 2019-08-23 17:32:05 -07:00
Philippe Tillet
f98b0b8e2a [general] deleted the old compiler frontend 2019-08-23 17:28:02 -07:00
Philippe Tillet
8798d240dc matmul test passes 2019-08-23 17:13:30 -07:00
Philippe Tillet
64a6910644 [lang][parser] better support for attributes 2019-08-22 21:02:38 -07:00
Philippe Tillet
845c0e5b93 adding tunable parameters 2019-08-22 19:21:01 -07:00
Philippe Tillet
87072203c1 [codegen] triton-ir code generation does not crash 2019-08-22 17:27:10 -07:00
Philippe Tillet
a6ec807223 more debugging 2019-08-21 21:53:41 -07:00
Philippe Tillet
a23225ad37 more progress 2019-08-21 18:27:02 -07:00
Philippe Tillet
5224bbbe06 preparing codegen 2019-08-20 18:06:30 -07:00
Philippe Tillet
61f25f90eb basic parsing doesn't throw error 2019-08-20 16:22:43 -07:00
Philippe Tillet
bc11e31419 [lang] more progress on parser 2019-08-19 20:56:39 -07:00
Philippe Tillet
0970fe12dd [general] cleaned tensorflow source code generation 2019-08-18 15:39:36 -07:00
Philippe Tillet
457c330f15 more cleaning 2019-08-18 14:20:42 -07:00
Philippe Tillet
c787ebae68 more cleaning 2019-08-18 14:09:55 -07:00
Philippe Tillet
81571246cf [general] fixed some warnings 2019-08-18 14:08:57 -07:00
Philippe Tillet
c05445d001 [general] removed dnn/ module and runtime/jit.cpp 2019-08-18 00:41:05 -07:00
Philippe Tillet
b58b0d8b27 [general] removed unnecessary includes 2019-08-18 00:34:30 -07:00
Philippe Tillet
b4a9ed9663 [python] added basic tensorflow support 2019-08-17 18:18:26 -07:00
Philippe Tillet
078f0052fe more cleaning 2019-08-17 16:12:17 -07:00
Philippe Tillet
11a6a92598 [python][tensorflow] basic op generation is working 2019-08-16 20:50:18 -07:00
Philippe Tillet
c7cb5f82ad [general] removed LLVM #include's in all Triton headers 2019-08-16 15:56:58 -07:00
Philippe Tillet
4de22df930 [python] added skeleton for python interface 2019-08-15 20:50:10 -07:00
Philippe Tillet
3ece461ce2 added tensorflow code generator 2019-08-15 15:59:53 -07:00
Philippe Tillet
38a8b0ab19 [runtime] overall of the run-time API 2019-08-14 20:26:11 -07:00
Philippe Tillet
b8cd63e0da [codegen] separated lower_dot_inst into lower_outer_dot ||
lower_hmma_dot || lower_scanline_dot
2019-08-12 21:48:30 -07:00
Philippe Tillet
4bc5758a22 [general] some cleaning:
* trans/dot -> peephole
* isel -> added function for tile-level lowering
2019-08-12 21:15:21 -07:00
Philippe Tillet
1400d960a6 [auto-tuning] much smaller parameters space 2019-08-12 21:15:21 -07:00