triton

Author	SHA1	Message	Date
Keren Zhou	16aed94ff5	[Analysis/Allocation] Allocation passes now assumes that slices always alias (#108 ) This code in this branch assumes the `src` operand in `insert_slice_async` always aliases the result, which shouldn't hold for generally cases but is just a workaround to make the pipeline pass work. I'm also working on the complete analysis in another [branch](https://github.com/openai/triton-mlir/tree/keren/analyze-slice).	2022-09-09 12:03:41 -07:00
Keren Zhou	328b87aec6	Keren/tensor slice insert alloc (#94 ) This branch defines three new triton_gpu operations to partially solve #87. Below is an overview: ``` %tensor = triton_gpu.alloc_tensor : tensor<2x16x16xf16, #A> %b = triton_gpu.insert_slice_async %a_ptr, %tensor, %offset {axis = 0 : i32, cache = 1 : i32, evict = 1 : i32, isVolatile = false} : tensor<16x16x!tt.ptr<f16>, #AL> -> tensor<2x16x16xf16, #A> %c = triton_gpu.extract_slice %b, %offset {axis = 0 : i32} : tensor<2x16x16xf16, #A> -> tensor<16x16xf16, #A> ``` We plan to fully replace `copy_async` with `insert_slice_async`. This hasn't been done yet.	2022-09-01 12:37:17 -07:00
Shintaro Iwasaki	84aa7d025a	[TritonIR] simplify Load/StoreOps when mask is true/false (#79 ) * [TritonIR] fix Load/Store/CopyAsyncOp's parsers * [TritonIR] simplify Load/StoreOps when mask is true/false * [TEST] adds tests to check load/store simplification	2022-08-24 12:55:49 -07:00
Philippe Tillet	192be76b3c	[OPTIMIZER] Rewrite patterns for layout conversions (#64 )	2022-08-18 12:49:37 -07:00
Shintaro Iwasaki	2ba9a83465	[BUILD] fix minor issues with MLIR assert enabled (#46 )	2022-08-11 21:20:47 -07:00
Philippe Tillet	d1593e6ca8	[TritonGPU] Improved documentation and semantics of layout encodings (#30 )	2022-07-31 13:59:44 -07:00
Philippe Tillet	432c3df265	[BUILD] MacOS can now build compiler and run MLIR tests (#25 )	2022-07-27 01:32:10 -07:00
Philippe Tillet	6d62d88d4f	[CI] run clang-format (#24 )	2022-07-26 17:25:03 -07:00
Keren Zhou	96cc6fb563	[TritonGPU] Pretty printer for layouts (#21 )	2022-07-26 10:50:11 -07:00
Yan Da	63e6a85901	Fix blocked layout parser	2022-07-15 15:19:11 +08:00
Yan Da	9d1b5e3f79	special encoding for broadcast	2022-06-18 21:16:45 +08:00
Yan Da	7b09b5f9e9	the pipeline pass now generates and accepts valid IR	2022-06-07 19:34:59 +08:00
Yan Da	366dddc3bc	update mma encoding & triton-opt	2022-06-06 21:03:58 +08:00
Yan Da	7807f64ef3	rename sharded_layout => blocked_layout	2022-06-05 16:14:59 +08:00
Yan Da	d5eca56cf3	more TritonGPU unit tests	2022-06-05 14:25:09 +08:00
Da Yan	e36a54eb86	more progress on the definition of layouts	2022-05-31 11:43:21 +00:00
Yan Da	441fd7c3cc	assembly format	2022-05-25 17:53:24 +08:00
Yan Da	9b670cfb9f	Add ReduceOp	2022-05-25 14:15:36 +08:00
Yan Da	a2c9f919a8	TritonGPU verifier	2022-05-24 19:48:56 +08:00
Yan Da	96876a46d1	More progress on Triton=>TritonGPU conversion (works for matmul)	2022-05-09 21:19:53 +08:00
Yan Da	3ad7bee35e	More conversion patterns	2022-05-04 12:50:02 +08:00
Yan Da	75d32e2442	More on TritonGPU conversion	2022-05-02 21:51:00 +08:00
Yan Da	1428185c9c	More progress on TritonGPUTypeConverter & TritonGPUConversionTarget	2022-05-01 22:06:54 +08:00
Yan Da	2239ac1998	more progress on TritonGPU	2022-04-28 18:51:31 +08:00

24 Commits