blender

Author	SHA1	Message	Date
Brecht Van Lommel	21a535840d	Fix T53270: crash with multiscatter GGX after recent refactoring. In fact this was an existing issue when exceeding the number of available closure, but it's more common now that we set the number to 0 for shadows and emission	2017-11-09 20:28:00 +01:00
Mai Lavelle	087331c495	Cycles: Replace __MAX_CLOSURE__ build option with runtime integrator variable Goal is to reduce OpenCL kernel recompilations. Currently viewport renders are still set to use 64 closures as this seems to be faster and we don't want to cause a performance regression there. Needs to be investigated. Reviewed By: brecht Differential Revision: https://developer.blender.org/D2775	2017-11-09 01:04:06 -05:00
Brecht Van Lommel	26f39e6359	Cycles: add bevel shader, for raytrace based rounded edges. The algorithm averages normals from nearby surfaces. It uses the same sampling strategy as BSSRDFs, casting rays along the normal and two orthogonal axes, and combining the samples with MIS. The main concern here is that we are introducing raytracing inside shader evaluation, which could be quite bad for GPU performance and stack memory usage. In practice it doesn't seem so bad though. Note that using this feature can easily slow down renders 20%, and that if you care about performance then it's better to use a bevel modifier. Mainly this is useful for baking, and for cases where the mesh topology makes it difficult for the bevel modifier to work well. Differential Revision: https://developer.blender.org/D2803	2017-11-07 22:35:12 +01:00
Brecht Van Lommel	f79f386731	Code refactor: rename subsurface to local traversal, for reuse.	2017-11-07 22:35:12 +01:00
Brecht Van Lommel	d0af56fe3b	Cycles: antialias normal baking if the mesh has a bump map.	2017-11-07 22:35:12 +01:00
Brecht Van Lommel	e74b229342	Fix incorrect MIS weights in Cycles with multiple lights. This causes some difference in the classroom scene, where ray visibility tricks are used and break the MIS balance. Otherwise there doesn't seem to be much effect, but better to use the right formulas. Problem originally identified by Lukas.	2017-11-07 22:35:12 +01:00
Brecht Van Lommel	8a72be7697	Cycles: reduce closure memory usage for emission/shadow shader data. With a Titan Xp, reduces path trace local memory from 1092MB to 840MB. Benchmark performance was within 1% with both RX 480 and Titan Xp. Original patch was implemented by Sergey. Differential Revision: https://developer.blender.org/D2249	2017-11-05 20:48:33 +01:00
Brecht Van Lommel	c571be4e05	Code refactor: sum transparent and absorption weights outside closures.	2017-11-05 18:13:44 +01:00
Brecht Van Lommel	2c02a04c46	Code refactor: remove emission and background closures, sum directly.	2017-11-05 18:13:44 +01:00
Brecht Van Lommel	cac3d4d166	Cycles: fix inefficient attribute map storage, saves 615MB in victor scene.	2017-11-05 18:00:48 +01:00
Brecht Van Lommel	5801ef71e4	Code refactor: device memory cleanups, preparing for mapped host memory.	2017-11-05 15:22:04 +01:00
Sergey Sharybin	71f46bc367	Cycles: Add utility function to distinguish between scatter and absorption volume ID	2017-11-01 11:10:51 +01:00
Sergey Sharybin	5d7138c08a	Cycles: Cleanup, make it more obvious what preprocessor belongs to	2017-11-01 11:10:10 +01:00
Sergey Sharybin	7f45acee80	Cycles: Cleanup, delete trailing whitespace	2017-11-01 11:06:55 +01:00
Brecht Van Lommel	bbc7eb8ae5	Cycles: restore SOBOL_SKIP hack, for some cases where it helps still.	2017-10-29 16:44:20 +01:00
Brecht Van Lommel	171c4e982f	Cycles: use AO factor to let user adjust intensity of AO bounces. We are already using the AO distance, so might as well offer this extra control over the intensity. Useful when an interior scene is supposed to be significantly darker than the background shader.	2017-10-25 21:46:23 +02:00
Brecht Van Lommel	7ad9333fad	Code refactor: store device/interp/extension/type in each device_memory.	2017-10-24 01:03:59 +02:00
Brecht Van Lommel	d85a0a722e	Fix part of T53038: principled BSDF clearcoat weight has no effect with 0 roughness.	2017-10-18 23:35:54 +02:00
Campbell Barton	99520e3f92	Cleanup: use 'e' prefix for enum typedefs Convention was only followed loosely, apply to DNA where changes aren't likely to conflict. (Skipped ModifierType for eg).	2017-10-17 13:49:20 +11:00
Brecht Van Lommel	2e50add164	Fix OpenCL performance regression after cubic interpolation. Reorganize code to reduce register pressure.	2017-10-15 17:46:50 +02:00
Sergey Sharybin	8d73ba58b6	Cycles: Fix compilation of sm_20 and sm_21 kernels Was broken since the bicubic commit for GPU support.	2017-10-10 12:26:02 +05:00
Brecht Van Lommel	cdb0b3b1dc	Code refactor: use DeviceInfo to enable QBVH and decoupled volume shading.	2017-10-08 13:17:33 +02:00
Brecht Van Lommel	f61c340bc1	Cycles: OpenCL bicubic and tricubic texture interpolation support.	2017-10-08 02:55:44 +02:00
Brecht Van Lommel	c040dedc12	Fix incorrect MIS with principled BSDF and specular roughness 0.	2017-10-07 22:10:02 +02:00
Brecht Van Lommel	d7eabc6765	Code cleanup: simplify cmake kernel install.	2017-10-07 15:32:20 +02:00
Brecht Van Lommel	2d92988f6b	Cycles: CUDA bicubic and tricubic texture interpolation support. While cubic interpolation is quite expensive on the CPU compared to linear interpolation, the difference on the GPU is quite small.	2017-10-07 15:30:57 +02:00
Brecht Van Lommel	23098cda99	Code refactor: make texture code more consistent between devices. * Use common TextureInfo struct for all devices, except CUDA fermi. * Move image sampling code to kernels//kernel__image.h files. * Use arrays for data textures on Fermi too, so device_vector<Struct> works.	2017-10-07 14:53:14 +02:00
Sergey Sharybin	837383ac78	Cycles: Cleanup, indendation	2017-10-06 19:33:59 +05:00
Sergey Sharybin	a950af8e24	Fix T53012: Shadow catcher creates artifacts on contact area The issue was caused by light sample being evaluated to nan at some point. This is root of the cause which is to be fixed, but is very hard to trace down especially via ssh (the issue only happens on AVX2 release build). Will give it a closer look when back to my AVX2 machine. For until then this is a good check to have anyway, it corresponds to what's happening in regular radiance sum.	2017-10-06 17:27:34 +05:00
Sergey Sharybin	0d3c8d0701	Cycles: Cleanup, indentation and wrapping	2017-10-06 16:54:37 +05:00
Brecht Van Lommel	4537e85584	Fix T53001: more workarounds for crash in AMD compiler with recent drivers.	2017-10-05 17:57:58 +02:00
Brecht Van Lommel	fb99ea79f8	Code refactor: split displace/background into separate kernels, remove luma.	2017-10-05 17:57:58 +02:00
Brecht Van Lommel	6da6f8d33f	Cycles: CUDA faster rendering of small tiles, using multiple samples like OpenCL. The work size is still very conservative, and this doesn't help for progressive refine. For that we will need to render multiple tiles at the same time. But this should already help for denoising renders that require too much memory with big tiles, and just generally soften the performance dropoff with small tiles. Differential Revision: https://developer.blender.org/D2856	2017-10-04 21:58:47 +02:00
Brecht Van Lommel	77f300e2a9	Fix use of uninitialized memory in Cycles normal baking.	2017-10-04 21:11:14 +02:00
Brecht Van Lommel	5bb677e592	Code refactor: zero render buffers outside of kernel. This was originally done with the first sample in the kernel for better performance, but it doesn't work anymore with atomics. Any benefit was very minor anyway, too small to measure it seems.	2017-10-04 21:11:14 +02:00
Brecht Van Lommel	12f4538205	Code refactor: use split variance calculation for mega kernels too. There is no significant difference in denoised benchmark scenes and denoising ctests, so might as well make it all consistent.	2017-10-04 21:11:14 +02:00
Brecht Van Lommel	e3e16cecc4	Code refactor: remove rng_state buffer and compute hash on the fly. A little faster on some benchmark scenes, a little slower on others, seems about performance neutral on average and saves a little memory.	2017-10-04 21:11:14 +02:00
Brecht Van Lommel	5b7d6ea54b	Code refactor: add WorkTile struct for passing work to kernel. This makes sharing some code between mega/split in following commits a bit easier, and also paves the way for rendering multiple tiles later.	2017-10-04 21:11:14 +02:00
Brecht Van Lommel	660e8e59e7	Fix T52645, T52645: AMD OpenCL compiler crash with recent drivers. Work around the bug by reshuffling code.	2017-10-04 21:00:46 +02:00
Brecht Van Lommel	f55735e533	CMake: support CUDA 9 toolkit, and automatically disable sm_2x binaries. Fermi cards (GTX 4xx and 5xx) are no longer supported with this version, so we can keep supporting both CUDA 8 and 9 for a while.	2017-10-01 14:14:53 +02:00
Brecht Van Lommel	d2bbd41b4e	Fix Cycles OpenCL compiler error after recent changes.	2017-09-29 14:54:10 +02:00
Brecht Van Lommel	400e6f37b8	Cycles: reduce subsurface stack memory usage. This is done by storing only a subset of PathRadiance, and by storing direct light immediately in the main PathRadiance. Saves about 10% of CUDA stack memory, and simplifies subsurface indirect ray code.	2017-09-28 15:18:43 +02:00
Sergey Sharybin	cb6f07f59e	Cycles: Cleanup, indentation	2017-09-25 11:15:54 +05:00
Sergey Sharybin	c0480bc972	Cycles: Fix compilation error of OpenCL megakernel on Apple	2017-09-23 17:07:19 +05:00
Sergey Sharybin	b460b8fb4a	Cycles: Fix compilation error of megakernel on NVidia device It is more readable to explicitly compare to NULL anyway.	2017-09-23 17:03:02 +05:00
Brecht Van Lommel	07ec0effb6	Code cleanup: simplify kernel side work stealing code.	2017-09-21 22:29:18 +02:00
Brecht Van Lommel	90d4b823d7	Cycles: use defensive sampling for picking BSDFs and BSSRDFs. For the first bounce we now give each BSDF or BSSRDF a minimum sample weight, which helps reduce noise for a typical case where you have a glossy BSDF with a small weight due to Fresnel, but not necessarily small contribution relative to a diffuse or transmission BSDF below. We can probably find a better heuristic that also enables this on further bounces, for example when looking through a perfect mirror, but I wasn't able to find a robust one so far.	2017-09-20 19:38:08 +02:00
Brecht Van Lommel	095a01a73a	Cycles: slightly improve BSDF sample stratification for path tracing. Similar to what we did for area lights previously, this should help preserve stratification when using multiple BSDFs in theory. Improvements are not easily noticeable in practice though, because the number of BSDFs is usually low. Still nice to eliminate one sampling dimension.	2017-09-20 19:38:08 +02:00
Brecht Van Lommel	b3afc8917c	Code cleanup: refactor BSSRDF closure sampling, for next commit.	2017-09-20 19:38:08 +02:00
Brecht Van Lommel	d029399e6b	Code cleanup: remove SOBOL_SKIP hack, seems no longer needed.	2017-09-20 19:38:08 +02:00

1 2 3 4 5 ...

1948 Commits